InTowards AIbyAmpatishan Sivalingam·May 22Forcing 16×16 Patches into 4096D Text Manifolds: A Mechanistic Model of Multimodal Projection…The quiet architectural crime at the heart of every vision language model you have shippedA response icon1A response icon1
InTowards AIbyAmpatishan Sivalingam·May 21Your Edge LLM is Memory Bound: Trading Compute for Bandwidth to Hit 30 Tokens per Second via LiteRT…The Promise Nobody Warned You AboutA response icon1A response icon1
InTowards AIbyAmpatishan Sivalingam·May 21Tokenizing the Continuous: How Patch-Based Architectures Unlocked Zero-Shot Time Series at 200M…Why the pointwise prediction era failed, and how a single architectural insight changed everything about how we forecast at scaleA response icon1A response icon1
InTowards AIbyAmpatishan Sivalingam·May 20Rendering for Latent Space: Why Agent Native Browsers Swapped Pixels for Tokenized Accessibility…How the browser engineering community stopped building for human eyes and started building for transformer weightsA response icon4A response icon4
InTowards AIbyAmpatishan Sivalingam·May 18Benign Overparameterization: Deconstructing Why 100B Parameter Transformers Defy the Bias Variance…The rule said bigger models must overfit. Then we built GPT and it forgot to read the rule.
InTowards AIbyAmpatishan Sivalingam·May 18Trading Zeros for Geometry: How Reshaping Transformer Weights to 2:4 Structured Sparsity Halves…We spent the last five years chasing ghosts inside our weight matrices. And for a long time, we were convinced we were winning.A response icon1A response icon1
InData Science CollectivebyAmpatishan Sivalingam·May 17Evicting Trillion-Parameter APIs: The Inference Tradeoffs Driving Pinterest’s Pivot to Self-Hosted…How the industry’s most expensive prototyping crutch became a production liability — and why small, sharp models running on your own…A response icon2A response icon2
InData Science CollectivebyAmpatishan Sivalingam·May 17Test Time Compute Over Parameter Scaling: How Hermes Agent’s Self Correction Loop Captured…The frontier just moved. And it did not move by adding more weights.
InTowards AIbyAmpatishan Sivalingam·May 15Dot Products as Routing Engines: Deconstructing Matrix Projections in O(N²) Transformer AttentionUnderstanding how transformers actually route information through massive sequences requires abandoning simplistic analogies and diving…
InTowards AIbyAmpatishan Sivalingam·May 10Stop Flushing the KV Cache: How GitHub Trades VRAM for Compute to Cut Agentic Workflow Costs by 10xThe Era of Stateless Agents: Building Intelligence with Goldfish Memory