|
1.I Tried to Run Qwen3.8–27B on a 16GB Mac Mini with AirLLM. Here’s Exactly Where It Breaks(pub.towardsai.net)
39 pts·Towards AI··infra
2.Pew study confirms sharp rise of AI-written text on the web since ChatGPT's launch(the-decoder.com)
44 pts·The Decoder··research
39 pts·Simon Willison··tools
58 pts·arXiv cs.AI··research
5.Knowing but Not Saying: Preventing Factual Access Failures in LLM SFT via Recall-Anchored Distillation(arxiv.org)
69 pts·arXiv cs.AI··research
87 pts·arXiv cs.AI··research
74 pts·arXiv cs.AI··research
8.RouteScan: A Non-Intrusive Approach to Auditing MoE LLMs Safety via Expert Routing Telemetry(arxiv.org)
89 pts·arXiv cs.CL··research
46 pts·Towards AI··tools
79 pts·arXiv cs.AI··research
11.When Do LLMs Replace Fine-Tuned NLU? A Decision Framework for Intent Detection in Production Conversational Systems(arxiv.org)
38 pts·arXiv cs.AI··research
40 pts·arXiv cs.AI··research
13.AgentMercury: Your Agent Can Synthesize Verifiable Environments for Business Scenarios at scale(arxiv.org)
71 pts·arXiv cs.AI··research
76 pts·arXiv cs.AI··research
38 pts·arXiv cs.AI··research
48 pts·arXiv cs.AI··research
17.The Metanym Game: An LLM Benchmark Without Ground Truth That Rises With the Models It Measures(arxiv.org)
42 pts·arXiv cs.AI··research
76 pts·arXiv cs.LG··research
68 pts·arXiv cs.LG··release
51 pts·arXiv cs.LG··research
71 pts·arXiv cs.AI··research
78 pts·arXiv cs.AI··research
23.TH-GNN: Heterogeneous Temporal Graph Neural Networks for LLM-Agent Shilling Attack Detection(arxiv.org)
69 pts·arXiv cs.CL··research
24.Nexus: Depth-Adaptive KV-Cache Splicing and Retrieval-Decoupled Tool Routing for Agentic LLMs on Unified Memory(arxiv.org)
69 pts·arXiv cs.AI··research
25.Certified Multi-Turn Robustness for LLM Safety via Compositional Bounds and Safety Persistence(arxiv.org)
85 pts·arXiv cs.AI··research
30 pts·arXiv cs.AI··research
27.Auditable by Construction: An Ontology-Driven Framework for Trustworthy LLM Analytics in Enterprise Finance(arxiv.org)
54 pts·arXiv cs.AI··research
28.Calibrating Criterion Revision in LLM Agents: Failure Modes and a Trace-Anchored Protocol(arxiv.org)
89 pts·arXiv cs.AI··research
34 pts·arXiv cs.AI··research
44 pts·arXiv cs.AI··research
30 of 500