Check the Fridge Before You Wake Gemma
This is a follow-up to Small Models Decide, Code Does, LLMs Think. In that post I argued that the model is only a third of the answer: small models decide, code does, and LLMs think, only when they have to. It was about taking model calls out of a browser agent. This time I wanted the same split in the most common LLM interface there is: a chat box. Most chat messages don’t need a large language model....
Small Models Decide, Code Does, LLMs Think
I spent an evening taking model calls out of a browser agent. A Google Flights task went from 15 seconds to under 7. The biggest gains came from moving dates, ranking, and field values into code. That experiment changed how I think about where models belong in a system. What prompted it On September 15, TypeSafe AI launched Jev and called it a System One model, after Kahneman’s fast, intuitive System 1....
बोल — Shipping Marathi TTS in Your Browser
I shipped a 326 MB Marathi text-to-speech model that runs entirely in your browser via WebGPU — no server, no API keys, just <audio> and a phoneme tokenizer in TypeScript. Type Marathi (or Marathi mixed with English — Minglish), pick a voice, hit synthesize. Try it yourself: huggingface.co/spaces/shreyask/bol-tts-marathi. Requires WebGPU (Chrome 113+, Edge 113+). The model is ~326 MB, loads once, then cached in your browser. First load takes ~30s; subsequent visits are instant....
I Gave an AI a Vocabulary of Only Emoji — And It Started to Communicate
In AMC’s Pantheon, a scientist named David Kim is uploaded — his mind digitized, his body gone. His daughter discovers he’s still there, trapped inside a computer, when he starts sending her emoji through her phone. He can think in full sentences. He can remember. He can love. But the only channel he has to reach her is the emoji keyboard. I couldn’t stop thinking about that scene. So I built it....
Take Back Your Feed: Building a Privacy-First ML-Powered Content Filter
Every feed you use — Hacker News, Reddit, X — has the same problem. Either it’s chronological and you’re drowning in noise, or it’s ranked by an algorithm optimizing for engagement, not your actual interests. You scroll past 90% of what you see, hoping something relevant catches your eye before the dopamine loop wins. The usual solutions don’t help much. RSS gives you firehose-level control but no ranking. Keyword filters are blunt — blocking “crypto” also hides legitimate cryptography research....