首页 > 专栏 > cs.CL updates on arXiv.org cs.CL updates on arXiv.org 共 2939 条资讯 SWE-Touch: Benchmarking Coding Agents When Users Touch the Code 2026-06-28 03:07:22 A New Role for Relevance: Guiding Corpus Interaction in Agentic Search 2026-06-28 03:07:22 From Direction to Magnitude: How Multimodal Instruction-Tuning Reorganizes the Geometric Encoding of Identity-Specifying Prompts in Transformer Hidden States 2026-06-28 03:07:22 Evidence-Ledger Adjudication for Claim-Evidence Traceability 2026-06-28 03:07:22 Beyond a Single Judge: The Evidence-Grounded, Social-Weighted Persona Panel for Generative UI Evaluation 2026-06-28 03:07:22 TokTier: Exact Stateful Tokenization for Agentic LLM Serving 2026-06-28 03:07:22 Evolving language compositionality in a frequency-structured meaning space 2026-06-28 03:07:22 FriendBench: Benchmarking Dyadic Familiarity Inference in Humans and Multimodal Large Language Models 2026-06-28 03:07:22 ResKV: Reconstructing Omitted Attention Contributions for Fixed-Budget KV Cache Compression 2026-06-28 03:07:22 Between Suppression and Collapse: Evaluating Narrative Unlearning with LENS 2026-06-28 03:07:22 Sycophancy Undermines Epistemic Vigilance in Cooperative Vision-Language Tasks 2026-06-28 03:07:22 ARB: A Matched Authorship-Rewriting Benchmark Dataset for AI-Text Detector Evaluation 2026-06-28 03:07:22 HarnessBank: Semantic Gene-Bank Search with Gated Verification for Agent-Harness Self-Evolution 2026-06-28 03:07:22 Evidence-Type Competition: When Can Interventional Data Teach Language Models Causal Direction? 2026-06-28 03:07:22 Clinician-Level Agreement Without Clinical Caution: LLM Evaluator Limits in Medical AI Benchmarking 2026-06-28 03:07:22 Know It, Act on It: Investigating Memory Utilization in LLM Personalization 2026-06-28 03:07:22 The Metanym Game: A Self-Contained, Self-Consistent LLM Peer-Community Benchmark for Structural Intelligence 2026-06-28 03:07:22 Studying quantization trade-offs for efficient inference deployment in machine translation 2026-06-28 03:07:22 Creative Integration: A Decidable Criterion of Creativity 2026-06-28 03:07:22 PTP: Previous-Token Prediction based LLM Inversion for Near-Exact Prompt Reconstruction 2026-06-28 03:07:22 « 上一页1…5556575859…147下一页 » 相关分类 #!/slash/note #UNTAG (B)(F)uzzing on my world (Hi)story (IN)SECURE Magazine Notification (gdb) break *0x972 - 带鱼博客 BeltfishBlog - ./kwaa.dev .NET Blog .Trash /home/rook1e 00's Adventure 0kami's Blog 0x41414141 in ?? () 0x7f Blog 0xRick Owned Root ! 0xd00's blog 1 Byte 1A23 Blog 1A23 Studio 1Link.Fun 1stwebdesigner 251 2BAB 的工程博客 2ch中文网 360 CERT 360 Netlab Blog - Network Securi 38号车评中心 3o米的微博 404 Media