StartupWiki is a community-driven AI-powered research directory of global startup ventures. Explore in-depth profiles of AI, Biotech, CleanTech, FinTech, Cybersecurity, and other deep-tech startups with funding data, financial metrics, competitive analysis, and team insights.
Browse categories: AI & Machine Learning, FinTech, Biotech, Cybersecurity, CleanTech, Quantum Computing, and more.
Browse all startups A-Z for the complete directory.
Read our blog for startup insights and deep dives.
The LPU Inference Engine — fastest AI inference at 1/10th the energy
HQ: Mountain View, CA, United States | Founded: 2016 | Employees: 250-500 | Stage: Late Stage VC | Website: https://groq.com
Groq is an AI hardware and cloud inference company founded in 2016 by Jonathan Ross — the Google engineer who invented the Tensor Processing Unit (TPU) — along with Douglas Wightman (former X lab) and a team from the original TPU program. Headquartered in Mountain View, California, Groq developed the Language Processing Unit (LPU) Inference Engine based on its proprietary Tensor Streaming Processor (TSP) architecture. Unlike GPUs that rely on parallel cores and dynamic runtime scheduling, Groq's single-core, deterministic design with ~230MB on-chip SRAM and ~80 TB/s bandwidth executes workloads scheduled entirely at compile time, eliminating cache misses and pipeline stalls to deliver predictable, ultra-low latency inference for large language models such as Meta Llama 3.1, Mistral Mixtral
Browse: Home | Blog | About | All Startups A-Z | View full profile on StartupWiki