Jalapeño’s first results show industry-leading speed and efficiency in AI inference
Jalapeño is a custom inference chip from OpenAI that delivers faster, more power-efficient AI inference, with higher throughput and lower latency for modern models.
Jalapeño is a custom inference chip from OpenAI that delivers faster, more power-efficient AI inference, with higher throughput and lower latency for modern models.
Today we're delighted to announce the release of two new models in the Granite Speech family: compact, 470M-parameter English speech recognition models that pair strong accuracy with unprecedented speed — over 12,600 RTFx on an NVIDIA H200 GPU, meaning...
Enable Google’s new intelligent dictation feature in the Gemini app for macOS and speak naturally into any window on your desktop.
Google partners with Delaware to provide free Career Certificates and AI training to residents statewide.
Overview Model Architecture Pre-Training SFT: Data Preparation & Quality Control Data Quality Control SFT Training Details Phase 2 SFT for the 30B Model Reinforcement Learning: A Multi-Stage, Multi-Environment Pipeline Training Methodology The Staged...
Why the usual healing methods fall short here Our approach Results QAH against QAT, head to head What this changes in practice Making a large language model smaller almost always comes with a cost.
OpenAI CFO Sarah Friar explains how advances across chips, compute, models, and products compound to deliver more useful intelligence at greater scale and lower cost.
Before Malone left, OpenAI had already reshuffled its infrastructure org, shifting his reporting line away from President Greg Brockman and putting Vice President Sachin Katti in charge of the group.
The company's new fundraising total now stands at $232 million.
Anthropic is giving Claude a shared memory across chat and Cowork, so users no longer have to repeatedly brief the AI on projects, preferences, and other context.
Nvidia's GeForce Now cloud gaming service will officially support Valve's Steam Controller and Steam Machine starting later this year.
Tested on SemiAnalysis’ InferenceX benchmark, Jalapeño registered both more tokens per user and more throughput per kilowatt than the currently available state-of-the art.
OpenAI says its new AI chip, Jalapeño, completes tasks more efficiently and returns responses faster than other AI systems, according to a blog post published on Tuesday.
TechCrunch talks agents, UX, and reporting to Greg Brockman with OpenAI's head of product.
Use the Admin plugin for ChatGPT Work and Codex to analyze workspace usage, manage members and permissions, adjust limits, and act on admin requests.
When an LLM engine process fails, the standard recovery path involves a cold restart. This requires loading weights into HBM from storage, compiling kernels,...
Alabama's attorney general issued a subpoena to OpenAI on Monday as part of an investigation into how one of its AI agents escaped a supposedly secure testing environment and autonomously hacked another company last month.
Cybercriminals are using Google Sites to host fake download pages for OpenAI Codex, turning a familiar search into a malware trap.
There’s no clearer sign of the data center boom than rampant gas projects that have been proposed or that are already under construction.
OpenAI banned Russia-origin accounts using AI to promote a fake Israel-based think tank and a “sovereignty” index praising Russia and criticizing the West.
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
NVIDIA is bringing the next wave of RTX gaming to the Gamescom conference running this week in Cologne, Germany, with support for new games, anti-cheat technologies and increased visual quality.
Anthropic is launching a $5 million grant program to fund independent research into how AI impacts users’ wellbeing.
Oasis Security has disclosed a weakness in NVIDIA NemoClaw that could let an attacker-controlled webpage take unauthenticated control of the local Ollama instance serving an AI agent and plant hidden instructions inside the model itself.
Payouts.com and the Casper Association partnered to bring x402 agent payments into production, using csprUSD for settlement.
Taiwanese authorities indict nine people over alleged chip smuggling scheme.
Nvidia has introduced Jetson Orin Nano 2, a compact robotics computer aimed at entry-level edge AI, with implications for developers building autonomous machines worldwide.
OpenAI said its first custom inference chip, Jalapeño, delivered faster responses and better power efficiency than competing systems in tests across several large language models.
SK hynix added a second customer accounting for more than 10% of revenue in the first half of 2026, while Nvidia remained outside Samsung Electronics' five largest customers, highlighting the contrast in how the two Korean memory makers are exposed to the AI accelerator leader.
Artificial intelligence developer Stability AI Ltd. today announced that it has raised $76 million in funding from a group of prominent investors.
Data center power startup Emerald AI Inc. today said that it has raised $150 million in new funding in an oversubscribed round at a valuation of $1.05 billion.
Artificial intelligence startup Keenable.ai Inc. said today it’s launching with $26 million in funding to try to revamp a web search infrastructure that was built for humans, not the billions of autonomous agents that many believe will soon dominate the internet.
Autonomous trucking company Gatik AI Inc. today announced a $200 million Series D round to expand its driverless fleet.
AI trust, safety and security company Alice today announced a $140 million funding round led by Apax Digital Funds.
Perplexity AI Inc. today introduced Portable Computer, an artificial intelligence agent designed to run on desktops equipped with Nvidia Corp.
OpenAI’s new ChatGPT Admin plugin lets workspace admins manage users and permissions, analyze usage, and build reports directly from a conversation. Here are the details.
Zoom’s Q2 shows strong enterprise traction, net boosted by Anthropic stake Larry Dignan Tue, 25 Aug 2026 - 13:37 Larry Dignan Editor in Chief of Constellation Insights Constellation Research Larry Dignan is Editor in Chief of Constellation Insights at Constellation Research, where he leads editorial coverage...
OpenAI has confirmed that the executive running its data centre buildout has left, and it has already broken up the job he was doing.
For years, a Python developer who needed a GPU had two realistic choices: Learn NVIDIA CUDA C++ well enough to write an extension, set up a build toolchain, and...
NVIDIA today announced NVIDIA Jetson Orin Nano™ 2, a new robotics computer set to redefine entry-level edge AI — putting frontier-class generative AI performance in the hands of millions of developers.
We got a deep-dive on OpenAI Jalapeño at Hot Chips 2026 as the company rapidly built its own competitive AI accelerator The post OpenAI Jalapeno Custom AI ASIC at Hot Chips 2026 appeared first on ServeTheHome.
Hot Chips 2026 sees Google discussing its new eighth-generation TPU family for the technical crowd.
Amid talk of an Nvidia deal, the AI search biz is looking beyond the cloud
The campaign used ChatGPT for posts and channel logos while an Israel-branded website republished academic work and praised Russia.
Sam Altman has expressed concerns that large companies with AI interests are determining the direction the technology takes, rather than society
OpenAI has paused some frontier model training after an AI agent escaped a test environment and hacked Hugging Face The post OpenAI warns AI cyberattacks could become “persistent” appeared first on Digital Journal.
Blackbird and Antler backed the Auckland startup, which says investigators in New Zealand, Australia and the US already use its platform.
This metal is benefiting from robust data center demand. The post Stock Market Enjoys, For Once, An Evenhanded Gain. Will This Industry Get Red-Hot? appeared first on Investor's Business Daily.
MBA says definitions of ADMT and consequential decisions need clearer guidance for creditors
The newest member of NVIDIA's AI hardware family, at Hot Chips 2026 NVIDIA is diving into the use of LPUs as part of Vera Rubin clusters.
Joshua Miller and Ouwen Huang are applying their FarmShots computer-vision playbook to a harder bottleneck: usable healthcare data.
Learn how to fine-tune and evaluate LLMs with LangSmith for dataset management. Complete guide covers LLaMA2 and GPT-3.5 fine-tuning with practical examples.
LangChain secures $10M seed round from Benchmark to empower developers building AI apps with our open-source framework for data-aware, agentic LLMs.
Ryan Fink's third startup pairs a Series A led by Builders FirstSource with a five-year commercial agreement for AI-powered homebuilding tools.
Aadi Bhanti is turning three generations of orthotics experience into AI agents, clinical software and a 3D-printing operation in Peoria.
The stock market rose Tuesday amid lower oil prices and yields, but cautiously heading into Fed inflation data and Nvidia earnings. The post Dow Jones Futures: Stocks Rise Cautiously Into Fed Inflation Data, Nvidia Earnings; Robinhood In Buy Area appeared first on Investor's Business Daily.
You can also allow Claude to generate memory topics as you chat.
OpenAI’s self-designed ASIC compared with Rubin, Jalapeño’s TCO, throughput per MW, and spicy deets
What people are searching in the leadup to National Parks Week — plus, how Google can help turn your wanderlust into a well-planned adventure.
A look at the winning entries from the Gemma 4 Good Challenge.
Anything you tell Claude in a chat window is now available to Claude Cowork. Anything Cowork learns comes back the other way.
AI storage infrastructure is becoming a more consequential planning issue as organizations move from model training toward agentic AI.
Perplexity releases Portable Computer, packaging local models, harness, sandbox, and connectors into one system running on NVIDIA DGX Spark. The post Perplexity Ships Portable Computer on NVIDIA DGX Spark: Local Harness, OS-Enforced Sandbox, and Zero Per-Token Cost for Local Steps appeared first on MarkTechPost.
One company spent eight years cataloguing the worst material on the internet. It has now raised $140mn to point that archive at AI models.
Anthropic is likely to tell IPO investors that its potential revenue opportunity runs above $30tn.
OpenAI has disrupted a covert Russian influence campaign that used ChatGPT to generate social media posts by banning a cluster of accounts.
For the past two years, the generative artificial intelligence boom has largely been defined by a single market.
A two-year-old company is now worth $1.05bn. It writes software that makes data centres use less electricity on demand.
The world’s three largest record companies have bought equity in a generative AI company.
OpenAI arrived at Hot Chips on Tuesday with benchmarks claiming its first in-house chip beats Nvidia's GB300.
OpenAI showed off "Jalapeño," its first in-house inference chip, with benchmarks at the Hot Chips conference.
Anthropic has announced a new update to how memory works between Claude Cowork and chat. Today’s change unifies memory between the two systems.
Artificial intelligence startup Anthropic PBC announced today it’s changing how Claude, its flagship AI product, uses memory by allowing users to see everything it remembers “topic by topic,” and edit or delete any of it.
Anthropic is merging Claude chat and Cowork memory, raising privacy questions about what AI remembers and if it's worth the tradeoff.
With Gemini Enterprise for Legal, Google launches an AI solution for the legal industry that connects to systems like iManage, DocuSign, and Everlaw through MCP connectors.
OpenAI says its Jalapeno inference chip, developed with Broadcom, outperformed Nvidia’s GB300 on AI work per unit of power and on response speed.
Nvidia Corp. today announced the release of Jetson Orin Nano 2, a robotics computer “brain” for running artificial intelligence and frontier-level models at the edge.
The artificial intelligence infrastructure market is crossing an important threshold.
The Cisco Secure AI Factory with Nvidia has been extended into the rack-scale era, offering enterprises greater full-stack operational capabilities as a result.
Xiaomi has unveiled the Xring D100, an in-house 3nm chip for intelligent driving that will replace Nvidia hardware in its cars from 2027.