- Published on
Exploring how 1B-parameter small language models can outperform 405B giants through clever compute allocation, test-time scaling, and innovative training strategies. Is bigger really always better in the AI world?
7 min read
Read more →I'm Rohith, a Research Scientist at the Centre for Development of Telematics (C-DOT), and a deep learning enthusiast with a passion for peeling back the layers of multimodal machine learning systems. This blog is intended to make recent advances easier to grasp, and is an open invitation to think deeper, explore further, and stay curious about the evolving landscapes of AI.
Enjoyed what you read? I'd love to hear your thoughts—feel free drop a comment on any article. For updates on new posts, Subscribe to the newsletter and get the latest straight to your inbox 📬.
Showing 35 of 35 posts