LLM-as-Jev: LLMs Are Already Jev-Style Decision Models -- When and How to Fine-Tune Them Paper • 2610.02076 • Published 3 days ago • 7
orcarouter/OrcaSAQ-2-Cyber-27B-Uncensored-GGUF Text Generation • 27B • Updated 5 days ago • 18.1k • 437
Don't Mask the Environment: Observation Supervision Changes How Agents Explore Under RL Paper • 2609.20715 • Published 20 days ago • 44
VC-Attention: Value Smoothing and Softmax Casting for Low-bit Attention Paper • 2609.15810 • Published 23 days ago • 51
Continual Learning Mechanisms Compose for Long-Horizon Memorization Paper • 2609.06986 • Published about 1 month ago • 376
CORE: Improving Compositional Reasoning in MLLM Embedding via Reranker Distillation Paper • 2609.04083 • Published Sep 3 • 27
MemTrapBench: Benchmarking Cognitive Traps in LLM Memory Use Paper • 2608.20202 • Published Aug 20 • 34
OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution Paper • 2608.00677 • Published Aug 1 • 266
SkillZip: Evaluation-Free Skill Compression for Self-Evolving Agents by Discovering Reusable Structure Paper • 2608.11079 • Published Aug 11 • 17