Hub
AI & Machine Learning
How modern AI systems are built, evaluated and made efficient, without the hype.
This week
- AI & ML Inference efficiencySpeculative Decoding Under Real Serving LoadSpeculative decoding promises faster text generation. This preprint measures what happens once real traffic, batching and memory limits enter the picture.
9:21 - AI & ML Reinforcement learningTeaching Agents to Explore When Rewards Are RareA simulation study tests whether a simple curiosity bonus can help reinforcement learning agents find rewards that appear only once in a long while.
9:02 - AI & ML LLMsDo Long Context Windows Actually Get Used?Models advertise windows of a million tokens. A new benchmark asks how much of that window they can really use when it counts.
9:46
Topics in this hub
LLMsReinforcement learningEvaluationInference efficiency
Every new AI & Machine Learning paper, the day it's out.
Coming soon to the App Store