view article Article Meta is back with Muse Glimmer: local, agentic, multimodal, and open source +2 pcuenq, merve, burtenshaw, ariG23498 • about 16 hours ago • 37
MerchantBench: Benchmarking LLM Agents for Long-Term Coherence in E-Commerce Operations Paper • 2607.28956 • Published 11 days ago • 96
view article Article Welcome Inkling by Thinking Machines +3 burtenshaw, merve, pcuenq, ariG23498, andito • 27 days ago • 160
view article Article Designing the hf CLI as an agent-optimized way to work with the Hub celinah, Wauplin • Jun 4 • 60
view article Article ITBench-AA: Frontier Models Score Below 50% on the First Benchmark for Agentic Enterprise IT Tasks — by Artificial Analysis and IBM ibm-research • May 27 • 18
view article Article Profiling in PyTorch (Part 1): A Beginner's Guide to torch.profiler +3 ariG23498, sayakpaul, sergiopaniego, ror, pcuenq • May 29 • 158
view article Article Beyond LLMs: Why Scalable Enterprise AI Adoption Depends on Agent Logic ibm-research • Jun 1 • 89
📝 Research & Long-Form Blog Posts Collection In-depth technical articles and research pieces published by Hugging Face • 20 items • Updated 4 days ago • 35
view article Article Open Responses: What you need to know +2 evalstate, burtenshaw, merve, pcuenq • Jan 15 • 112
view article Article Liberate your OpenClaw +6 clem, burtenshaw, pcuenq, jeffboudier, merve, nielsr, victor, mishig • Mar 27 • 49
view article Article Harness, Scaffold, and the AI Agent Terms Worth Getting Right sergiopaniego, ariG23498 • May 25 • 137
view article Article DeepSeek-V4: a million-token context that agents can actually use burtenshaw • Apr 24 • 53