Trending...
- California: Governor Newsom signs new law to protect workers, require disclosures on AI-generated advertising - 165
- California: Governor Newsom expands film and TV tax credits with new legislation, creates tax credit to support post-production jobs - 117
- California: Governor Newsom issues legislative update 9.18.26 - 113
Cortex adds context, prompt prefix and exact response reuse, delivering up to 150× faster prefill, 11× faster response delivery and 2.25× throughput.
SAN FRANCISCO - Californer -- Pervaziv AI today announced a new 3-Tier Cortex Inference Cache Architecture designed to make repeated enterprise AI work faster while preserving the controls required for current, authorized and trustworthy results.
The architecture introduces three forms of reuse across Cortex: context reuse, prompt prefix reuse and exact response reuse. Each tier has its own validity rules and security boundary.
"The next step in enterprise AI performance is not simply caching more," said Anoop Jaishankar, Founder and CEO of Pervaziv AI. "It is knowing exactly what can be reused, what changed, who is still authorized to use it, and when fresh computation is required. Cortex is turning inference caching into a governed capability, where speed comes from removing repeated work without removing the checks that make the result trustworthy."
More on The Californer
Reuse What Is Safe. Recompute What Changed.
The first tier, context reuse, can avoid rebuilding eligible application context when source state and permissions remain valid. The model still reasons over the current request.
The second tier, prompt prefix reuse, can reuse eligible model preparation when the beginning of a model request remains identical. In repeated tests, prompt processing fell from 2,913 milliseconds to 19.3 milliseconds for a tested workload, approximately 150 times faster, while the model continued generating a new response.
The third tier, exact response reuse, is designed for narrowly approved, identical, read only requests. In one live test, an initial request completed in 3,132 milliseconds and an identical repeat returned a reported cache hit in 281 milliseconds, approximately 11 times faster with about 91 percent lower latency.
A separate workload matrix completed 240 requests with zero request failures and showed up to 2.25 times the throughput under moderate concurrency for one mid length workload, while also reinforcing the need to evaluate tail latency and capacity alongside speed.
More on The Californer
The architecture follows recent Cortex releases that expanded where enterprise AI work can happen. Cortex Connect established continuity, Cortex Cloud added durable managed execution, and Cortex Discover brought Cortex into a dedicated agentic AI browser. The new architecture focuses on repeated work across those experiences.
Cortex evaluates reuse against authorization, source freshness, model and tool configuration, task state and route eligibility. When reuse cannot be proven valid, it falls back to fresh computation.
About Pervaziv AI
Pervaziv AI builds Cortex, an Enterprise AI Control Layer designed to coordinate specialized AI models, agents, search, skills, security, privacy, verification and governed execution across browser, mobile, development and cloud environments. The company focuses on helping organizations move from AI assistance toward trusted, controlled outcomes. Learn more at https://pervaziv.com.
The architecture introduces three forms of reuse across Cortex: context reuse, prompt prefix reuse and exact response reuse. Each tier has its own validity rules and security boundary.
"The next step in enterprise AI performance is not simply caching more," said Anoop Jaishankar, Founder and CEO of Pervaziv AI. "It is knowing exactly what can be reused, what changed, who is still authorized to use it, and when fresh computation is required. Cortex is turning inference caching into a governed capability, where speed comes from removing repeated work without removing the checks that make the result trustworthy."
More on The Californer
- Lunai Bioworks (N A S D A Q: LNAI) Takes Parkinson's Discovery to the Next Level With Exclusive Tanaist Agreement
- OakBloomIQ Launches Free AI Visibility Score, Reveals How AI Rates You
- Revenue Optics Names Prat Patibandla Director of Growth Marketing
- California: Governor Newsom delivers $886 million in utility bill relief, with millions of households receiving an average of $75 this summer
- Independent Autopsies Are Changing Civil Cases: Kansas City Forensic Explains Why Families Are Seeking Second Opinions
Reuse What Is Safe. Recompute What Changed.
The first tier, context reuse, can avoid rebuilding eligible application context when source state and permissions remain valid. The model still reasons over the current request.
The second tier, prompt prefix reuse, can reuse eligible model preparation when the beginning of a model request remains identical. In repeated tests, prompt processing fell from 2,913 milliseconds to 19.3 milliseconds for a tested workload, approximately 150 times faster, while the model continued generating a new response.
The third tier, exact response reuse, is designed for narrowly approved, identical, read only requests. In one live test, an initial request completed in 3,132 milliseconds and an identical repeat returned a reported cache hit in 281 milliseconds, approximately 11 times faster with about 91 percent lower latency.
A separate workload matrix completed 240 requests with zero request failures and showed up to 2.25 times the throughput under moderate concurrency for one mid length workload, while also reinforcing the need to evaluate tail latency and capacity alongside speed.
More on The Californer
- JEGS Launches Transformed Digital Commerce Platform Powered by PhaseZero
- Subject matter experts: Present your ideas at IISE Annual Conference 2027 in Louisville
- xBxBio Files U.S. Provisional Patent Application for xBxBio's Virtual Heart™
- Salt Security Extends Its Agentic Security Platform with Native AI Detection and Response
- Cummings Graduate Institute Celebrates New Doctor of Behavioral Health Graduates
The architecture follows recent Cortex releases that expanded where enterprise AI work can happen. Cortex Connect established continuity, Cortex Cloud added durable managed execution, and Cortex Discover brought Cortex into a dedicated agentic AI browser. The new architecture focuses on repeated work across those experiences.
Cortex evaluates reuse against authorization, source freshness, model and tool configuration, task state and route eligibility. When reuse cannot be proven valid, it falls back to fresh computation.
About Pervaziv AI
Pervaziv AI builds Cortex, an Enterprise AI Control Layer designed to coordinate specialized AI models, agents, search, skills, security, privacy, verification and governed execution across browser, mobile, development and cloud environments. The company focuses on helping organizations move from AI assistance toward trusted, controlled outcomes. Learn more at https://pervaziv.com.
Source: Pervaziv AI
0 Comments
Latest on The Californer
- Ahead of Climate Week NYC, Governor Newsom announces California cut climate pollution again as economy keeps growing
- Scale Your Side Hustle Overnight: Why Smart Creators Are Ditching Traditional Livestreaming
- California: Governor Newsom signs most comprehensive data center laws in the nation, providing communities more control on water, electricity, and land use
- Explosive New Book Warns Toxic Peril to Black and Brown Communities in America
- Bank Statement Loans Up to $30 Million: Lendmire Announces High-Net-Worth Financing
- Marc Yaffee Returns to Rolling Hills Casino Saturday, October 3
- Governor Newsom proclaims state of emergency to bolster statewide El Nino preparedness, protect California
- NYC Big Book Award Celebrates 10th Anniversary with Announcement of 2026 Winners
- Things to Do in Temecula This Weekend: Free Wellness Event Sept. 26
- Countrywide Rental Provides Reliable Portable Restroom Solutions in Bynum, Alabama
- Repair Shop Solutions Launches Most Comprehensive Photo Editing for Digital Vehicle Inspections
- California: Governor Newsom issues legislative update 9.20.2026
- 12 Simple Practices Proven to Switch on Your Body's "Happy Genes"
- California establishes Dolly Parton Day
- California: Governor Newsom signs bills as Los Angeles readies for 2028 Olympic and Paralympic Games
- Governor Gavin Newsom signs legislation to accelerate California's EV future, expand consumer choice, and strengthen energy independence
- Best Dodge Lemon Law Attorney in Los Angeles County
- Care En Route Launches Location-Aware Booking For Mobile Healthcare Practices
- What they are saying: California leaders praise Governor Newsom's democracy defending action ahead of November elections
- Governor Newsom signs bill expanding fuel options to help Californians save at the pump, amid Donald Trump's gas price crisis