Popular on EntSun
- Sea Tales Strengthens U.S. Leadership Team as Retail Expansion Accelerates - 444
- Stitt Entertainment Names Tristan Xavier Lead Stylist - 374
- ChargeOn Announces Conversational Payment Management Solution - 372
- Mecca to Mars: Earth's Escape? - 280
- Robert DeMaio, Phinge CEO to Speak at 30th IIPLA 2026 on Securing IP, User Data & Digital Sovereignty via Verified App-less Tech. Attend the Live Q&A! - 263
- Rehearsals Begin on 'The Call List' by Brian James Polak - 250
- Free JavaScript Charting Library ProEssentialsJS Permits Commercial Use. Highcharts, SciChart and LightningChart Free Tiers Do Not - 249
- Volant Announces Major Partnership with the STEM Racing Canada Program to CHAMPION YOUTH ENGINEERING - 238
- Big Panic Records Announces "African Sunshine" by Believe, Shirley Jones & Omasta - 228
- Yalil Guerra's "Clave" to Receive World Premiere in New Symphonic Version at The Soraya - 211
Similar on EntSun
- PropAccount.com Launches Marketing Hub Within PropGenie, Giving Prop Firm Operators an Automated Marketing Team
- Solid Earth's new "Welcome to Real Estate" campaign introduces RealtyID™: Know Before Your Open the Door
- Lunai Bioworks (N A S D A Q: LNAI) Takes Parkinson's Discovery to the Next Level With Exclusive Tanaist Agreement
- Revenue Optics Names Prat Patibandla Director of Growth Marketing
- JEGS Launches Transformed Digital Commerce Platform Powered by PhaseZero
- Subject matter experts: Present your ideas at IISE Annual Conference 2027 in Louisville
- Put Your Herd on Your Phone: Kiko Nation Makes Livestock Management Simple
- Future Intelligence Think Tank Surpasses 1,000 Members Exploring Human and Artificial Intelligence
- International Society of Medical AI Convenes Global Faculty in Florence for ISMAI 2026
- Retell AI White Label Platform for Agencies Launched by VoiceAIWrapper, With Branded Client Portals and No Per-Minute Markup
Pervaziv AI Advances Cortex with 3-Tier Inference Cache Architecture for Faster, Trusted AI Workflows
EntSun News/11103152
Cortex adds context, prompt prefix and exact response reuse, delivering up to 150× faster prefill, 11× faster response delivery and 2.25× throughput.
SAN FRANCISCO - EntSun -- Pervaziv AI today announced a new 3-Tier Cortex Inference Cache Architecture designed to make repeated enterprise AI work faster while preserving the controls required for current, authorized and trustworthy results.
The architecture introduces three forms of reuse across Cortex: context reuse, prompt prefix reuse and exact response reuse. Each tier has its own validity rules and security boundary.
"The next step in enterprise AI performance is not simply caching more," said Anoop Jaishankar, Founder and CEO of Pervaziv AI. "It is knowing exactly what can be reused, what changed, who is still authorized to use it, and when fresh computation is required. Cortex is turning inference caching into a governed capability, where speed comes from removing repeated work without removing the checks that make the result trustworthy."
More on EntSun News
Reuse What Is Safe. Recompute What Changed.
The first tier, context reuse, can avoid rebuilding eligible application context when source state and permissions remain valid. The model still reasons over the current request.
The second tier, prompt prefix reuse, can reuse eligible model preparation when the beginning of a model request remains identical. In repeated tests, prompt processing fell from 2,913 milliseconds to 19.3 milliseconds for a tested workload, approximately 150 times faster, while the model continued generating a new response.
The third tier, exact response reuse, is designed for narrowly approved, identical, read only requests. In one live test, an initial request completed in 3,132 milliseconds and an identical repeat returned a reported cache hit in 281 milliseconds, approximately 11 times faster with about 91 percent lower latency.
A separate workload matrix completed 240 requests with zero request failures and showed up to 2.25 times the throughput under moderate concurrency for one mid length workload, while also reinforcing the need to evaluate tail latency and capacity alongside speed.
More on EntSun News
The architecture follows recent Cortex releases that expanded where enterprise AI work can happen. Cortex Connect established continuity, Cortex Cloud added durable managed execution, and Cortex Discover brought Cortex into a dedicated agentic AI browser. The new architecture focuses on repeated work across those experiences.
Cortex evaluates reuse against authorization, source freshness, model and tool configuration, task state and route eligibility. When reuse cannot be proven valid, it falls back to fresh computation.
About Pervaziv AI
Pervaziv AI builds Cortex, an Enterprise AI Control Layer designed to coordinate specialized AI models, agents, search, skills, security, privacy, verification and governed execution across browser, mobile, development and cloud environments. The company focuses on helping organizations move from AI assistance toward trusted, controlled outcomes. Learn more at https://pervaziv.com.
The architecture introduces three forms of reuse across Cortex: context reuse, prompt prefix reuse and exact response reuse. Each tier has its own validity rules and security boundary.
"The next step in enterprise AI performance is not simply caching more," said Anoop Jaishankar, Founder and CEO of Pervaziv AI. "It is knowing exactly what can be reused, what changed, who is still authorized to use it, and when fresh computation is required. Cortex is turning inference caching into a governed capability, where speed comes from removing repeated work without removing the checks that make the result trustworthy."
More on EntSun News
- Drake Brings Jimmy Brooks Back as KayeDinero Turns Banks Ave Into a Nighttime Visual
- Transformational $104 Million Musculoskeletal Healthcare Opportunity as Expansion Strategy Accelerates for Cardiff Lexington Corp (Stock Symbol: CDIX)
- PropAccount.com Launches Marketing Hub Within PropGenie, Giving Prop Firm Operators an Automated Marketing Team
- Save 15 Percent Off Florida Keys Accommodations with KeysCaribbean's Advance Purchase Rate Discount
- Woodside Rehab & Nursing Enters a New Era of Clinical Excellence and Growth
Reuse What Is Safe. Recompute What Changed.
The first tier, context reuse, can avoid rebuilding eligible application context when source state and permissions remain valid. The model still reasons over the current request.
The second tier, prompt prefix reuse, can reuse eligible model preparation when the beginning of a model request remains identical. In repeated tests, prompt processing fell from 2,913 milliseconds to 19.3 milliseconds for a tested workload, approximately 150 times faster, while the model continued generating a new response.
The third tier, exact response reuse, is designed for narrowly approved, identical, read only requests. In one live test, an initial request completed in 3,132 milliseconds and an identical repeat returned a reported cache hit in 281 milliseconds, approximately 11 times faster with about 91 percent lower latency.
A separate workload matrix completed 240 requests with zero request failures and showed up to 2.25 times the throughput under moderate concurrency for one mid length workload, while also reinforcing the need to evaluate tail latency and capacity alongside speed.
More on EntSun News
- Thar Process Partners with French Leader ExtrateX to Bring Prep to Process-Scale SFC to Europe
- iPOP! Alum Lucas Till Stars Alongside Zach Braff in Upcoming Film Clean Hands
- Hispaniola Cigars Distributor Explores the Vodou Spirit Behind Its Baron Samedi Cigar in Educatio
- Port St. Lucie REALTOR® Irene Pernice Earns SFR® Certification
- AMR and UAS Take the Stage at WorldSkills Shanghai 2026
The architecture follows recent Cortex releases that expanded where enterprise AI work can happen. Cortex Connect established continuity, Cortex Cloud added durable managed execution, and Cortex Discover brought Cortex into a dedicated agentic AI browser. The new architecture focuses on repeated work across those experiences.
Cortex evaluates reuse against authorization, source freshness, model and tool configuration, task state and route eligibility. When reuse cannot be proven valid, it falls back to fresh computation.
About Pervaziv AI
Pervaziv AI builds Cortex, an Enterprise AI Control Layer designed to coordinate specialized AI models, agents, search, skills, security, privacy, verification and governed execution across browser, mobile, development and cloud environments. The company focuses on helping organizations move from AI assistance toward trusted, controlled outcomes. Learn more at https://pervaziv.com.
Source: Pervaziv AI
0 Comments
Latest on EntSun News
- ClearSight Therapeutics Receives NIH SBIR Phase I Award for Conjunctivitis and EKC Research
- Follow No One Entertainment Launches AI-Free Record Label & Global Broadcast Network
- Sidow Sobrino on Michael Jackson and the Art of Building an Identity of Your Own
- Lunai Bioworks (N A S D A Q: LNAI) Takes Parkinson's Discovery to the Next Level With Exclusive Tanaist Agreement
- Murder is on the Menu for Halloween Weekend at Georgia's Lanier Islands Resort
- The DMV Jazz Series returns to New Spire Arts with Five New Artists
- Revenue Optics Names Prat Patibandla Director of Growth Marketing
- Independent Autopsies Are Changing Civil Cases: Kansas City Forensic Explains Why Families Are Seeking Second Opinions
- JEGS Launches Transformed Digital Commerce Platform Powered by PhaseZero
- Subject matter experts: Present your ideas at IISE Annual Conference 2027 in Louisville
- Cummings Graduate Institute Celebrates New Doctor of Behavioral Health Graduates
- Visit London Appoints The SEO Works for AI-First Search Strategy
- King of the Wild Frontier Reading at North Coast Repertory Theatre
- For One-Hit Wonder Day: Treat yourself to Two Great Songs called "Big Bucks" and "Run For Office"
- Dual-Media Launches The Studio Lab™
- Inglewood Associates and Gordon Brothers Provide Update on Sale Process for 925 Euclid Avenue in Cleveland
- Century Fasteners Corp. Exhibiting at the 2026 International Fastener Expo
- Las Vegas Attorney Thomas Boley Publishes Free Plain-English Guide to 100 Nevada Criminal Laws, in English and Spanish
- Bank Statement Loans Up to $30 Million: Lendmire Announces High-Net-Worth Financing
- Marc Yaffee Returns to Rolling Hills Casino Saturday, October 3