arXiv Artificial IntelligenceStyle Wins, Substance Loses: A Diagnosis of LLM-as-Judge in Idea Generation
arXiv:2608.01666v2 Announce Type: replace-cross Abstract: However, whether these judges truly evaluate the…
Curated from international AI laboratories, specialist publications and technology outlets. Last update: 3 hours ago.
arXiv Artificial IntelligencearXiv:2608.01666v2 Announce Type: replace-cross Abstract: However, whether these judges truly evaluate the…
arXiv Artificial IntelligencearXiv:2608.01366v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are integral to…
arXiv Artificial IntelligencearXiv:2608.01269v2 Announce Type: replace-cross Abstract: Hierarchical Graph Retrieval-Augmented Generation…
arXiv Artificial IntelligencearXiv:2608.00747v2 Announce Type: replace-cross Abstract: Large language models are increasingly integrated…
arXiv Artificial IntelligencearXiv:2608.00151v2 Announce Type: replace-cross Abstract: Current evaluation frameworks for artificial…
arXiv Artificial IntelligencearXiv:2608.00123v2 Announce Type: replace-cross Abstract: LLM-native advertising embeds sponsored content…
arXiv Artificial IntelligencearXiv:2608.00076v2 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs)…
arXiv Artificial IntelligencearXiv:2608.00058v2 Announce Type: replace-cross Abstract: Accurate measurement of ECG intervals, including…
arXiv Artificial IntelligencearXiv:2607.28587v2 Announce Type: replace-cross Abstract: SWE-bench-like benchmarks are widely used for…
arXiv Artificial IntelligencearXiv:2607.27670v2 Announce Type: replace-cross Abstract: Jigsaw puzzle solving requires jointly reasoning…
arXiv Artificial IntelligencearXiv:2607.22999v2 Announce Type: replace-cross Abstract: Language agents can now interact fluently with…
arXiv Artificial IntelligencearXiv:2607.22739v2 Announce Type: replace-cross Abstract: We study how far a deliberately simple…