Document Retrieval-Aware Chunking (D-RAC): Universal Retrieval-Aware Ingestion of Enterprise Documents via PDF Normalization and Multimodal Markdown Conversion Paper • 2609.24220 • Published 11 days ago • 68
Grouped Value Attention: Efficient KV Caching via On-Demand Key Reconstruction Paper • 2609.13285 • Published 24 days ago • 81
Think Again or Think Longer? Selective Verification for Budget-Aware Reasoning Paper • 2606.19808 • Published Jun 18 • 3
Externalizing Research Synthesis and Validation in AI Scientists through a Research Harness Paper • 2606.18874 • Published Jun 17 • 7
Morpheus: A Morphology-Aware Neural Tokenizer and Word Embedder for Turkish Paper • 2606.18717 • Published Jun 17 • 5
Grouped Query Experts: Mixture-of-Experts on GQA Self-Attention Paper • 2606.20945 • Published Jun 18 • 81
Web Retrieval-Aware Chunking (W-RAC) for Efficient and Cost-Effective Retrieval-Augmented Generation Systems Paper • 2604.04936 • Published Jan 8 • 26
ZClip: Adaptive Spike Mitigation for LLM Pre-Training Paper • 2504.02507 • Published Apr 3, 2025 • 90
Beyond Extraction: Contextualising Tabular Data for Efficient Summarisation by Language Models Paper • 2401.02333 • Published Jan 4, 2024 • 7
Komodo: A Linguistic Expedition into Indonesia's Regional Languages Paper • 2403.09362 • Published Mar 14, 2024 • 11
Vision-Guided Chunking Is All You Need: Enhancing RAG with Multimodal Document Understanding Paper • 2506.16035 • Published Jun 19, 2025 • 89