首页 > 专栏 > cs.CL updates on arXiv.org cs.CL updates on arXiv.org 共 2754 条资讯 When Calibration Rankings Reverse: Accuracy-Controlled Evaluation for Fair Comparison of LLMs 2026-06-28 03:07:22 When transformers learn "impossible" languages, what do they learn? 2026-06-28 03:07:22 Multilingual Polarization Detection Using Transformer-Based Models with Class Weighting and Threshold Tuning 2026-06-28 03:07:22 Test-Time Verification for Text-to-SQL via Outcome Reward Models 2026-06-28 03:07:22 Beyond Clean Text: Evaluating Encoder and Decoder Robustness for Bangla Event Detection in Noisy Text 2026-06-28 03:07:22 Training Therapeutic Judges and Multi-Agent Systems for Human-Aligned Mental Health Support 2026-06-28 03:07:22 ShopX: A Foundation Model for Intent-to-Item Fulfillment in Agentic Shopping 2026-06-28 03:07:22 The Decomposition Is the Fingerprint: Per-Component Identity for Agent Skills 2026-06-28 03:07:22 InfiniteWeb: Scalable Web Environment Synthesis for GUI Agent Training 2026-06-28 03:07:22 QVal: Cheaply Evaluating Dense Supervision Signals for Long-Horizon LLM Agents 2026-06-28 03:07:22 Fund2Persona: A Framework for Building and Refining Financial Advisor Personas from Fund Disclosure Data 2026-06-28 03:07:22 BenGER: Benchmarking LLM Systems on Subsumption-Based Legal Reasoning in German Law 2026-06-28 03:07:22 ReplicatorBench: Benchmarking LLM Agents for Replicability in Social and Behavioral Sciences 2026-06-28 03:07:22 BlockPilot: Instance-Adaptive Policy Learning for Diffusion-based Speculative Decoding 2026-06-28 03:07:22 Beyond Scalar Rewards: Dense Feedback for LLM Policy Synthesis in Sequential Social Dilemmas 2026-06-28 03:07:22 MauBERT: Universal Phonetic Inductive Biases for Few-Shot Acoustic Units Discovery 2026-06-28 03:07:22 To Reason or to Fabricate: Reasoning Without Shortcuts via Hint-Anchored Pairwise Aggregation 2026-06-28 03:07:22 CLARity: Reasoning Consistency Alone Can Teach Reinforced Experts 2026-06-28 03:07:22 mamabench and mamaretrieval: Benchmarks for Evaluating Medical Retrieval-Augmented Generation in Maternal, Neonatal, and Reproductive Health 2026-06-28 03:07:22 Model Directions, Not Words: Mechanistic Topic Models Using Sparse Autoencoders 2026-06-28 03:07:22 « 上一页1…126127128129130…138下一页 » 相关分类 #!/slash/note #UNTAG (B)(F)uzzing on my world (Hi)story (IN)SECURE Magazine Notification (gdb) break *0x972 - 带鱼博客 BeltfishBlog - ./kwaa.dev .NET Blog .Trash /home/rook1e 00's Adventure 0kami's Blog 0x41414141 in ?? () 0x7f Blog 0xRick Owned Root ! 0xd00's blog 1 Byte 1A23 Blog 1A23 Studio 1Link.Fun 1stwebdesigner 251 2BAB 的工程博客 2ch中文网 360 CERT 360 Netlab Blog - Network Securi 38号车评中心 3o米的微博 404 Media