Multi-Turn Agentic Scientific Literature Search via Workflow Induction Paper • 2607.00597 • Published Jul 1 • 31
When RL Meets Adaptive Speculative Training: A Unified Training-Serving System Paper • 2602.06932 • Published Feb 6 • 1
SAW-INT4: System-Aware 4-Bit KV-Cache Quantization for Real-World LLM Serving Paper • 2604.19157 • Published Apr 21 • 2
OSCAR: Offline Spectral Covariance-Aware Rotation for 2-bit KV Cache Quantization Paper • 2605.17757 • Published May 18 • 66
Kitty: Accurate and Efficient 2-bit KV Cache Quantization with Dynamic Channel-wise Precision Boost Paper • 2511.18643 • Published Nov 23, 2025 • 1
Multi-Turn Agentic Scientific Literature Search via Workflow Induction Paper • 2607.00597 • Published Jul 1 • 31
UrduBench: An Urdu Reasoning Benchmark using Contextually Ensembled Translations with Human-in-the-Loop Paper • 2601.21000 • Published Jan 28 • 4
Improving Multilingual Capabilities with Cultural and Local Knowledge in Large Language Models While Enhancing Native Performance Paper • 2504.09753 • Published Apr 13, 2025 • 6
Robust and Fine-Grained Detection of AI Generated Texts Paper • 2504.11952 • Published Apr 16, 2025 • 12
Robust and Fine-Grained Detection of AI Generated Texts Paper • 2504.11952 • Published Apr 16, 2025 • 12
1-800-SHARED-TASKS @ NLU of Devanagari Script Languages: Detection of Language, Hate Speech, and Targets using LLMs Paper • 2411.06850 • Published Nov 11, 2024 • 3
Augmenting Legal Decision Support Systems with LLM-based NLI for Analyzing Social Media Evidence Paper • 2410.15990 • Published Oct 21, 2024 • 1