首页 > 专栏 > cs.CL updates on arXiv.org cs.CL updates on arXiv.org 共 2939 条资讯 Evaluating Federated Pre-Training: On the Reliability of Downstream Fine-Tuning and Intrinsic Evaluation 2026-06-28 03:07:22 The Formalism Trap: Are LLM-as-a-Judge Evaluators Blinded by Consensus Mimicry under Social Load? 2026-06-28 03:07:22 PEFT of SLM for Telecommunications Customer Support: A Comparative Study of LoRA Configurations with Energy Consumption Analysis 2026-06-28 03:07:22 Can Large Language Models Derive New Knowledge? A Dynamic Benchmark for Biological Knowledge Discovery 2026-06-28 03:07:22 Frontier AI performance across the business disciplines: a case-grounded benchmark of knowledge work and analytical reasoning 2026-06-28 03:07:22 Measuring Cognitive Engagement in Collaborative Discourse with an Extended ICAP Framework: Comparing Human Annotation, In-Context Learning, and Reflective LLM Agents 2026-06-28 03:07:22 The Self-Correction Illusion: Role Relabeling Gates Explicit Error Flagging in Large Language Models 2026-06-28 03:07:22 TORUS: A Test of Rendering-Understanding Self-Coherence for Unified Audio Models 2026-06-28 03:07:22 Language Models Agree With Each Other, Not With Readers 2026-06-28 03:07:22 Implicit Reasoning for Large Language Model-based Generative Recommendation 2026-06-28 03:07:22 FineInstructions: Scaling Synthetic Instructions to Pre-Training Scale 2026-06-28 03:07:22 Metareasoning constraints couple narratives, affect and cognition 2026-06-28 03:07:22 LLMs struggle to simulate human belief updates in controlled environments 2026-06-28 03:07:22 AISPA: User-Centric System Prompt Auditing for Large Language Model Applications 2026-06-28 03:07:22 Change2Task: From Repository Changes to Executable Coding Agent Tasks and Environments 2026-06-28 03:07:22 CACHE-UK: A Stability-Aware Memory Editor for Sequentially Updated Quantized LLMs in Finance 2026-06-28 03:07:22 TCA-SIR: Learning Target-Conditioned Abstractions for Scientific Inspiration Retrieval 2026-06-28 03:07:22 (Towards) Scalable Reliable Automated Evaluation with Large Language Models 2026-06-28 03:07:22 SVR: Self-Verifying Refinement via Joint Verdict-Confidence Reinforcement Learning for Adaptive Test-Time Compute 2026-06-28 03:07:22 MORFES: A Benchmark for Productive Inflectional Competence in Modern Greek 2026-06-28 03:07:22 « 上一页1…5960616263…147下一页 » 相关分类 #!/slash/note #UNTAG (B)(F)uzzing on my world (Hi)story (IN)SECURE Magazine Notification (gdb) break *0x972 - 带鱼博客 BeltfishBlog - ./kwaa.dev .NET Blog .Trash /home/rook1e 00's Adventure 0kami's Blog 0x41414141 in ?? () 0x7f Blog 0xRick Owned Root ! 0xd00's blog 1 Byte 1A23 Blog 1A23 Studio 1Link.Fun 1stwebdesigner 251 2BAB 的工程博客 2ch中文网 360 CERT 360 Netlab Blog - Network Securi 38号车评中心 3o米的微博 404 Media