首页 > 专栏 > cs.LG updates on arXiv.org cs.LG updates on arXiv.org 共 3022 条资讯 SATQuest: A Verifier for Logical Reasoning Evaluation and Reinforcement Fine-Tuning of LLMs 2026-06-28 03:08:51 SOS-LoRA: Static Orthogonal-Subspace Low-Rank Adaptation with Fixed Multi-Scale Scaling 2026-06-28 03:08:51 Learning Spatio-Temporal Foundation Models from Pure Synthetic Data 2026-06-28 03:08:51 Benchmarking Machine Learning Models for Multi-Omics-Based Breast Cancer Prediction 2026-06-28 03:08:51 A Systematic Investigation of RL-Jailbreaking in LLMs 2026-06-28 03:08:51 High-accuracy Low-Bit KV-Cache Quantization via Local Distribution Restoration 2026-06-28 03:08:51 Stochastic Resetting Accelerates Reinforcement Learning Beyond Random Search 2026-06-28 03:08:51 Self-Evolving Just-In-Time Memory for Proactive Embodied Safety 2026-06-28 03:08:51 When Benchmarks Lie: Evaluating Malicious Prompt Classifiers Under True Distribution Shift 2026-06-28 03:08:51 Let the Data Decide: Supervision Analysis, Capability Trade-offs, and Adaptive Objective Routing in Continued Pre-Training via Off-Policy Distillation 2026-06-28 03:08:51 Interpret Policies in Deep Reinforcement Learning using SILVER with RL-Guided Labeling: A Model-level Approach to High-dimensional and Multi-action Environments 2026-06-28 03:08:51 Learning Structural Manipulability in Gate-Level Netlists Using Graph Neural Networks 2026-06-28 03:08:51 Mediator: Memory-efficient LLM Merging with Less Parameter Conflicts and Uncertainty Based Routing 2026-06-28 03:08:51 CIGPO: Contextual Information-Gain Policy Optimization for Multi-Turn Evidence-Reading LLM Agents 2026-06-28 03:08:51 EVOLVE: Efficient Learned Volume Compression with Variable-Rate Encoding on a Cross-Domain Database 2026-06-28 03:08:51 RobustMAD: Evaluating Real-World Robustness of Multimodal Small Language Models for Deployable Anomaly Detection Assistants 2026-06-28 03:08:51 Entanglement geometry separates circuit cutting, classical hardness, and trainability 2026-06-28 03:08:51 TRACE: Trajectory-Based Safety Patch Learning for LLM Post-Training Realignment 2026-06-28 03:08:51 The Curvature Shadow: An Apparent Failure of Maximum-Entropy Equilibrium Selection is a Removable Artifact 2026-06-28 03:08:51 KernelBench-Verified: Do LLM-Generated Kernels Actually Beat PyTorch? 2026-06-28 03:08:51 « 上一页1…9394959697…152下一页 » 相关分类 #!/slash/note #UNTAG (B)(F)uzzing on my world (Hi)story (IN)SECURE Magazine Notification (gdb) break *0x972 - 带鱼博客 BeltfishBlog - ./kwaa.dev .NET Blog .Trash /home/rook1e 00's Adventure 0kami's Blog 0x41414141 in ?? () 0x7f Blog 0xRick Owned Root ! 0xd00's blog 1 Byte 1A23 Blog 1A23 Studio 1Link.Fun 1stwebdesigner 251 2BAB 的工程博客 2ch中文网 360 CERT 360 Netlab Blog - Network Securi 38号车评中心 3o米的微博 404 Media