The original source remains the canonical version for attribution, rights, and later updates.
Read original ↗Jalapeño’s first results show industry-leading speed and efficiency in AI inference
Jalapeño is a custom inference chip from OpenAI that delivers faster, more power-efficient AI inference, with higher throughput and lower latency for modern models.
Related stories
TechCrunch Founder Summit’s agenda revealed: Unlock fundraising, hiring, and AI insights in Boston on November 4
On November 4, TechCrunch’s Founder Summit will bring a vital one-day crash course on startup building to Boston’s SoWa Power Station. Founders shouldn’t have to learn the hardest lessons the hardest way, and this event is designed to make the challenges of starting a company easier and the highs that much greater. Instead of months of trial and error, you get direct access to the investors and founders who’ve already made the calls you’re now facing. We’re talking everything from fundraising and hiring to AI strategy, and topping your category. Don’t take our word for it. Explore the full agenda of speakers and sessions we have planned for Founder Summit below, and lock in your spot at the…
Snorkel AI triples valuation to $3.5B as demand for AI training data booms
Snorkel AI, a startup that helps AI labs and corporations build training datasets and simulated environments, has raised a $350 million Series E at a $3.5 billion valuation. The new round, which was led by Insight Partners and S32, valued the seven-year-old startup at nearly triple the $1.3 billion valuation it garnered when it raised $100 million in a Series D 17 months ago. Existing investors, including Addition, Lightspeed, Greylock, GV, and Wells Fargo, also participated in the round. While Snorkel originally provided software for data-labeling automation , it shifted last year to providing customers with completed datasets, an offering it calls data-as-a-service. Rather than operating p…
Better prompt caching for GPT-6
Learn how GPT-6 improves prompt caching with higher cache hit rates, new diagnostics, explicit breakpoints, and controls that reduce latency and costs.
