Popular on EntSun
- Sea Tales Strengthens U.S. Leadership Team as Retail Expansion Accelerates - 404
- Stitt Entertainment Names Tristan Xavier Lead Stylist - 335
- ChargeOn Announces Conversational Payment Management Solution - 332
- Mecca to Mars: Earth's Escape? - 239
- Phinge Exposes Massive AI Security Risks, Claiming Its Patented Hardware-Verified Architecture Is The Only Safeguard Against Surveillance Capitalism - 223
- Robert DeMaio, Phinge CEO to Speak at 30th IIPLA 2026 on Securing IP, User Data & Digital Sovereignty via Verified App-less Tech. Attend the Live Q&A! - 221
- Rehearsals Begin on 'The Call List' by Brian James Polak - 206
- Free JavaScript Charting Library ProEssentialsJS Permits Commercial Use. Highcharts, SciChart and LightningChart Free Tiers Do Not - 203
- Volant Announces Major Partnership with the STEM Racing Canada Program to CHAMPION YOUTH ENGINEERING - 193
- Big Panic Records Announces "African Sunshine" by Believe, Shirley Jones & Omasta - 189
Similar on EntSun
- 3ptechies Partners with Coolmuster to Give Away Data Recovery Software Licenses
- DuraFast Label Company Launches Seiko SLP850 2" Thermal Printer with Free Label Promotion
- PricZone Launches Online Shopping Platform Offering Electronics, Gaming, and More
- XRPPower Expands Platform Security With Enhanced Brand Protection and Official Verification Standards
- Bynn Intelligence Ranks #1 in NIST Child Online Safety Evaluation for Ages 13–16
- Joulescope JS320 Launches to Help Engineers Develop Battery-Powered Devices with Greater Confidence
eRacks Ships Benchmarked Dual Intel Arc Pro B70 AI Servers and Prices AI Provisioning at $1,495
EntSun News/11100842
Measured on the company's own bench: 54 tokens per second on Qwen3-14B, 26 on Qwen3.6-27B from a single card - and the provisioning behind those numbers is now a standard priced offering: $1,495, or $2,495 with a private RAG stack, included on flags
CAMPBELL, Calif. - EntSun -- eRacks Open Source Systems today published benchmark results measured on a dual Intel Arc Pro B70 server during its pre-ship provisioning pass, and announced that the provisioning work behind those numbers is now a standard priced offering across its AI server line.
The measured numbers, from the company's bench this week: Qwen3-14B generating 54 tokens per second, and the larger Qwen3.6-27B holding a sustained 26 tokens per second on a single B70 - faster than most people read. The serving stack is fully open source: llama.cpp's official Intel build in rootless Podman containers, exposing the industry-standard OpenAI-compatible API, with models resident entirely in GPU memory.
More on EntSun News
The hardware is the point. The Arc Pro B70 carries 32GB of VRAM (the GPU's onboard memory, the hard limit on what models fit) per card, so a two-card server fields 64GB of GPU memory for less than the list price of a single 96GB flagship datacenter card. For private AI, that ratio of memory to dollars is the value play of 2026.
"Benchmarks on a spec sheet are marketing. Benchmarks on your machine are engineering," said Joseph Wolff, founder and CTO of eRacks Systems. "Getting these cards to production took three fixes you will not find in any manual - GPU power management that puts cards to sleep permanently, container networking that resets every connection while the server looks healthy. We solved them on the bench, and every AI server we ship now leaves with its own measured numbers and the rebuild notes in the customer's hands."
That work is now a named product: eRacks AI Provisioning & Setup covers burn-in, GPU bring-up with every fix applied, deployment and benchmarking of the customer's chosen models on the customer's actual hardware, and full rebuild documentation - $1,495, or $2,495 including a private RAG stack (retrieval-augmented generation: chat plus a vector database answering from the customer's own documents, fully offline). It is included at no charge on flagship orders.
More on EntSun News
The line ships configured to order at live prices: eRacks/AIDAN with one B70 from $13,895, eRacks/AINSLEY with two B70s and 64GB of GPU memory - the configuration class benchmarked above - from $21,395, the full AI server line from $7,695, and the 8-GPU eRacks/HIGHLANDER flagship from $154,995. Configuration and the company's no-signup rent-versus-own calculator: https://eracks.com/products/ai-rackmount-servers/ and https://eracks.com/tco/
The measured numbers, from the company's bench this week: Qwen3-14B generating 54 tokens per second, and the larger Qwen3.6-27B holding a sustained 26 tokens per second on a single B70 - faster than most people read. The serving stack is fully open source: llama.cpp's official Intel build in rootless Podman containers, exposing the industry-standard OpenAI-compatible API, with models resident entirely in GPU memory.
More on EntSun News
- European Patent for ALS Program Expands the Story: HOPE Deploys Robotic TMS & FDA Commercialization Path Advances: NRx Pharmaceuticals: NAS DAQ: NRXP
- VICTURY Sports Announces Formation of Youth Sports League Built on the Official, Patented, and Award-Winning Keepy Uppy® Ball
- GOOBZ'D™ Turns "Did You Mean: Good?" Search Confusion Into a New Campaign
- The Podcast of Power Returns With Exclusive Bear McCreary Interview
- Berkshire Actor Charlotte Drablow Expands Screen Career
The hardware is the point. The Arc Pro B70 carries 32GB of VRAM (the GPU's onboard memory, the hard limit on what models fit) per card, so a two-card server fields 64GB of GPU memory for less than the list price of a single 96GB flagship datacenter card. For private AI, that ratio of memory to dollars is the value play of 2026.
"Benchmarks on a spec sheet are marketing. Benchmarks on your machine are engineering," said Joseph Wolff, founder and CTO of eRacks Systems. "Getting these cards to production took three fixes you will not find in any manual - GPU power management that puts cards to sleep permanently, container networking that resets every connection while the server looks healthy. We solved them on the bench, and every AI server we ship now leaves with its own measured numbers and the rebuild notes in the customer's hands."
That work is now a named product: eRacks AI Provisioning & Setup covers burn-in, GPU bring-up with every fix applied, deployment and benchmarking of the customer's chosen models on the customer's actual hardware, and full rebuild documentation - $1,495, or $2,495 including a private RAG stack (retrieval-augmented generation: chat plus a vector database answering from the customer's own documents, fully offline). It is included at no charge on flagship orders.
More on EntSun News
- GOOBZ'D™ Begins Official Character Rollout With CHOMPZ, Character #001
- Making Time for Mythic Adventure
- High Rise 6: Power to Escape Expands the Film Franchise With New Cartels, New Locations
- Congressman Chuck Edwards (NC-11) Presents ReadyCommunities Partnership 2026 National Service Award to Local Businessman / US Air Force Veteran
- Cress Creeks Farm Opens Fall Corn Maze and Baby Lamb Hayride Experience in Ellijay, Georgia
The line ships configured to order at live prices: eRacks/AIDAN with one B70 from $13,895, eRacks/AINSLEY with two B70s and 64GB of GPU memory - the configuration class benchmarked above - from $21,395, the full AI server line from $7,695, and the 8-GPU eRacks/HIGHLANDER flagship from $154,995. Configuration and the company's no-signup rent-versus-own calculator: https://eracks.com/products/ai-rackmount-servers/ and https://eracks.com/tco/
Source: eRacks Open Source Systems
0 Comments
Latest on EntSun News
- OneVizion Launches Vera AI, Bringing Connected Operational Context to Infrastructure Teams
- P-Wave Press Announces Pushing the Wave 2025 by L.A. Davenport
- DONGSHENG Titanium Recycling: Promoting a Closed-Loop Cycle for High-End Titanium Materials
- Antoine Morgan steps into the "MARRIAGE BOX" and Things are getting interesting!
- French Camp Academy Releases New Video Inviting Individuals to Consider a Life of Service
- LET US READ Documentary on Dyslexia and Literacy to Screen at TCL Chinese 6 in Hollywood
- Don Barnhart Brings His NFTG Tour Home as Delirious Celebrates One Year at Silver Sevens
- Defense & Space Strategy Strengthens as New Leadership Builds on NASA Results and Expanding Multi-Orbit Opportunities for Ascent Solar Technologies
- Suits & Shoot Returns for 3rd Annual Noir Affair
- 'The Last Howl' Documentary Advocates for Conservation and Wildlife Preservation
- Venice FF 2026, Samantha Casella screens The Serpent's Caress
- 8th Edition of Filming Italy Venice Awards Cinema Talent
- STS Capital Partners is pleased to announce the appointment of Barry Brown as Vice President, Business Development
- Qscription Technologies and NEOPATHOLOGY CORP. Sign MOU to Bring FDA-Cleared Lung Imaging AI into U.S. Clinical Practice
- Heritage at Manalapan Welcomes New Sales Team as Luxury Single-Family Home Community Continues to Grow
- Mandeville Pests May Pose Serious Health Risks for Your Family
- The Hot Sardines Return to Frederick for a Night of Vintage Jazz at the Weinberg Center
- Share your workplace safety solutions at 2027 Applied Ergonomics Conference
- Parents No Longer Have to Wait 3 Weeks for a Sleep Consultant: Nora Talks Tonight, Stays for 5 Days, Costs $89
- May The Worst Team Win! Loserball Kicks Off Another NFL Season of Hilarious Mayhem