arXiv Artificial IntelligenceOISD: On-Policy Internal Self-Distillation of Language Models
arXiv:2605.29089v2 Announce Type: replace-cross Abstract: Recent reinforcement learning (RL) post-training…
Curated from international AI laboratories, specialist publications and technology outlets. Last update: 3 hours ago.
arXiv Artificial IntelligencearXiv:2605.29089v2 Announce Type: replace-cross Abstract: Recent reinforcement learning (RL) post-training…
arXiv Artificial IntelligencearXiv:2605.28740v2 Announce Type: replace-cross Abstract: As large language models are increasingly deployed…
arXiv Artificial IntelligencearXiv:2605.28183v4 Announce Type: replace-cross Abstract: We introduce BenGER (Benchmark for German Law), a…
arXiv Artificial IntelligencearXiv:2605.28042v2 Announce Type: replace-cross Abstract: Modern large language models (LLMs) achieve…
arXiv Artificial IntelligencearXiv:2605.28006v2 Announce Type: replace-cross Abstract: Understanding how LLMs reason is hindered by a…
arXiv Artificial IntelligencearXiv:2605.27984v2 Announce Type: replace-cross Abstract: Speech language models (SpeechLMs) have achieved…
arXiv Artificial IntelligencearXiv:2605.27971v2 Announce Type: replace-cross Abstract: When large language models are fine-tuned to…
arXiv Artificial IntelligencearXiv:2605.27786v3 Announce Type: replace-cross Abstract: Large language models are known to contain…
arXiv Artificial IntelligencearXiv:2605.27480v3 Announce Type: replace-cross Abstract: Large language model (LLM) serving creates…
arXiv Artificial IntelligencearXiv:2605.27068v2 Announce Type: replace-cross Abstract: Social deduction games have become a popular…
arXiv Artificial IntelligencearXiv:2605.24614v2 Announce Type: replace-cross Abstract: Large language model (LLM) unlearning has emerged…
arXiv Artificial IntelligencearXiv:2605.22455v2 Announce Type: replace-cross Abstract: Real-world deployment of AI vision models is both…