首页 > 专栏 > cs.CL updates on arXiv.org cs.CL updates on arXiv.org 共 2661 条资讯 QuasiMoTTo: Quasi-Monte Carlo Test-Time Scaling 2026-06-28 03:07:22 Quantifying the Affective Gap: A Zero-Shot Evaluation of LLMs on Fine-Grained Emotion Taxonomies 2026-06-28 03:07:22 Persona Non Grata: LLM Persona-Driven Generations in MCQA are Unstable in Distinct Dimensions 2026-06-28 03:07:22 From Personas to Plot: Character-Grounded Multi-Agent Story Generation for Long-Form Narratives 2026-06-28 03:07:22 MindEdit-Bench: Benchmarking Object-Level Counterfactual Spatial Reasoning in VLMs from In-the-Wild Photos 2026-06-28 03:07:22 Beyond Document Grounding: Span-Level Hallucination Detection over Code, Tool Output, and Documents 2026-06-28 03:07:22 MolSafeEval: A Benchmark for Uncovering Safety Risks in AI-Generated Molecules 2026-06-28 03:07:22 MultiSynt/MT: Trillion-Token Multi-Parallel Pre-Training Data Translated Across 36 Languages 2026-06-28 03:07:22 When Classic Cache Policies Fail: Learning-Augmented Replacement for Semantic Retrieval Buffers 2026-06-28 03:07:22 How Ethos and Pathos Appeals Resonate in Reader Interpretations of Social Media Messages 2026-06-28 03:07:22 Watermarking for Proprietary Dataset Protection 2026-06-28 03:07:22 Dynamic Bidirectional Pattern Memory: A Production-Scale Empirical Characterisation of Inference-Time Gating in Clinical NLP 2026-06-28 03:07:22 Mapping the Evaluation Frontier: An Empirical Survey of the Bias-Reliability Tradeoff Across Eleven Evaluator-Agent Conditions 2026-06-28 03:07:22 CAT: Confidence-Adaptive Thinking for Efficient Reasoning of Large Reasoning Models 2026-06-28 03:07:22 Rosetta: Composable Native Multimodal Pretraining 2026-06-28 03:07:22 Recovering Input Text from Hidden States: Study of Gradient-Based Inversion of Decoder-Only Language Models 2026-06-28 03:07:22 Testing Frontier Large Language Models' Physics Literacy in Parallel Physical Worlds 2026-06-28 03:07:22 The Course of News Events: A Comparison of Bottom-Up and Top-Down Approaches for Collecting Text-Based Data about Disasters 2026-06-28 03:07:22 MetaHOPE: A Metaphor-Oriented Evaluation Framework for Analysing MT and LLM Translation Errors 2026-06-28 03:07:22 Measuring Reasoning Quality in LLMs: A Multi-Dimensional Behavioral Framework 2026-06-28 03:07:22 « 上一页1…113114115116117…134下一页 » 相关分类 #!/slash/note #UNTAG (B)(F)uzzing on my world (Hi)story (IN)SECURE Magazine Notification (gdb) break *0x972 - 带鱼博客 BeltfishBlog - ./kwaa.dev .NET Blog .Trash /home/rook1e 00's Adventure 0kami's Blog 0x41414141 in ?? () 0x7f Blog 0xRick Owned Root ! 0xd00's blog 1 Byte 1A23 Blog 1A23 Studio 1Link.Fun 1stwebdesigner 251 2BAB 的工程博客 2ch中文网 360 CERT 360 Netlab Blog - Network Securi 38号车评中心 3o米的微博 404 Media