I trained a small transformer in 1.5hrs and it beats many LLMs
Small transformer trained in 1.5 hours matches performance of many published language models.
“HN: 643 pts”
Why it matters
Weekly Silicon's editorial model scored this 2.45/10 overall, reading it above all as an AI-capability story — news that changes what models, and the labs behind them, can do. Its strongest dimension is technology breakthrough at 4/10 — a genuine capability or engineering advance rather than a routine product update — with regional relevance close behind at 3/10, pointing to direct impact on US technology hubs rather than a purely overseas development. The impact is global rather than tied to one US hub, so the thing to watch is how it filters into domestic supply chains and hiring.
Derived from the AI score breakdown below.
AI score breakdown
A composite of 2.45/10 put this story at #31 for Wednesday, September 2, 2026, driven mostly by technology breakthrough (4/10) and regional relevance (3/10).
Composite is the weighted sum of the five dimensions. How scoring works →
Related stories
- OpenAI’s next big AI model has ‘entered the AGI era’ · 2026-09-03
- Fei-Fei Li’s World Labs debuts Atlas, a world model showcase for advanced spatial intelligence · 2026-09-03
- Anthropic launches Claude Fable 5.1 after inking $35B cloud deal with Lambda · 2026-09-02
- OpenAI’s new reasoning technique alarms AI safety experts · 2026-09-02
- ~77% of all new "Success" self-help books on Amazon are likely written by AI, with 1 author, Noah Felix Bennett, publishing a stunning 74 books in mid-2025 alone, at a rate of >1 per day. Richard Trillion Mantey, who has published hundreds of books, was assessed to have used AI for every single book · 2026-04-09