Trending...
- Making Time for Mythic Adventure - 171
- Sky Quarry Enters a Powerful New Chapter: Refinery Restart Emminent, Nevada Oil Initiative and Visibility Put (NAS DAQ: SKYQ) in the Spotlight - 116
- Long Beach: City Recaps Preparedness and Response Efforts on Peninsula Following Hurricane Marie Impacts - 115
Measured on the company's own bench: 54 tokens per second on Qwen3-14B, 26 on Qwen3.6-27B from a single card - and the provisioning behind those numbers is now a standard priced offering: $1,495, or $2,495 with a private RAG stack, included on flags
CAMPBELL, Calif. - Californer -- eRacks Open Source Systems today published benchmark results measured on a dual Intel Arc Pro B70 server during its pre-ship provisioning pass, and announced that the provisioning work behind those numbers is now a standard priced offering across its AI server line.
The measured numbers, from the company's bench this week: Qwen3-14B generating 54 tokens per second, and the larger Qwen3.6-27B holding a sustained 26 tokens per second on a single B70 - faster than most people read. The serving stack is fully open source: llama.cpp's official Intel build in rootless Podman containers, exposing the industry-standard OpenAI-compatible API, with models resident entirely in GPU memory.
More on The Californer
The hardware is the point. The Arc Pro B70 carries 32GB of VRAM (the GPU's onboard memory, the hard limit on what models fit) per card, so a two-card server fields 64GB of GPU memory for less than the list price of a single 96GB flagship datacenter card. For private AI, that ratio of memory to dollars is the value play of 2026.
"Benchmarks on a spec sheet are marketing. Benchmarks on your machine are engineering," said Joseph Wolff, founder and CTO of eRacks Systems. "Getting these cards to production took three fixes you will not find in any manual - GPU power management that puts cards to sleep permanently, container networking that resets every connection while the server looks healthy. We solved them on the bench, and every AI server we ship now leaves with its own measured numbers and the rebuild notes in the customer's hands."
That work is now a named product: eRacks AI Provisioning & Setup covers burn-in, GPU bring-up with every fix applied, deployment and benchmarking of the customer's chosen models on the customer's actual hardware, and full rebuild documentation - $1,495, or $2,495 including a private RAG stack (retrieval-augmented generation: chat plus a vector database answering from the customer's own documents, fully offline). It is included at no charge on flagship orders.
More on The Californer
The line ships configured to order at live prices: eRacks/AIDAN with one B70 from $13,895, eRacks/AINSLEY with two B70s and 64GB of GPU memory - the configuration class benchmarked above - from $21,395, the full AI server line from $7,695, and the 8-GPU eRacks/HIGHLANDER flagship from $154,995. Configuration and the company's no-signup rent-versus-own calculator: https://eracks.com/products/ai-rackmount-servers/ and https://eracks.com/tco/
The measured numbers, from the company's bench this week: Qwen3-14B generating 54 tokens per second, and the larger Qwen3.6-27B holding a sustained 26 tokens per second on a single B70 - faster than most people read. The serving stack is fully open source: llama.cpp's official Intel build in rootless Podman containers, exposing the industry-standard OpenAI-compatible API, with models resident entirely in GPU memory.
More on The Californer
- Simplicity IT Named Finalist in 2026 MSP Titans of the Industry Awards for Community Impact
- Infinity Infusion Solutions Named Finalist for D Magazine's D CEO 2026 Excellence in Healthcare Awards
- Attn: Art Editors, Computer Programmers and of course, AI: Here is a painting about life in our computerized world. It's a painting called CLICK HERE
- California: Governor Newsom signs new law to protect workers, require disclosures on AI-generated advertising
- Pervaziv AI Unveils Cortex Discover, an Agentic AI Browser Built for Governed Enterprise Work
The hardware is the point. The Arc Pro B70 carries 32GB of VRAM (the GPU's onboard memory, the hard limit on what models fit) per card, so a two-card server fields 64GB of GPU memory for less than the list price of a single 96GB flagship datacenter card. For private AI, that ratio of memory to dollars is the value play of 2026.
"Benchmarks on a spec sheet are marketing. Benchmarks on your machine are engineering," said Joseph Wolff, founder and CTO of eRacks Systems. "Getting these cards to production took three fixes you will not find in any manual - GPU power management that puts cards to sleep permanently, container networking that resets every connection while the server looks healthy. We solved them on the bench, and every AI server we ship now leaves with its own measured numbers and the rebuild notes in the customer's hands."
That work is now a named product: eRacks AI Provisioning & Setup covers burn-in, GPU bring-up with every fix applied, deployment and benchmarking of the customer's chosen models on the customer's actual hardware, and full rebuild documentation - $1,495, or $2,495 including a private RAG stack (retrieval-augmented generation: chat plus a vector database answering from the customer's own documents, fully offline). It is included at no charge on flagship orders.
More on The Californer
- LovelySkin launches The Nine Aurora following the discovery of nine Skin Longevity Zones
- Governor Newsom visits Jay Leno's Garage to sign legislation celebrating and preserving California's classic car and lowrider heritage
- Mesa West Capital Originates $27 Million Loan for Seattle Apartment Acquisition
- Canadian Tint Expo 2027: Why Every Window Tinter Should Compete in Niagara Falls on April 3rd & 4th
- SkillFront Launches ISO 9001:2026 Enterprise Certification on September 16, 2026
The line ships configured to order at live prices: eRacks/AIDAN with one B70 from $13,895, eRacks/AINSLEY with two B70s and 64GB of GPU memory - the configuration class benchmarked above - from $21,395, the full AI server line from $7,695, and the 8-GPU eRacks/HIGHLANDER flagship from $154,995. Configuration and the company's no-signup rent-versus-own calculator: https://eracks.com/products/ai-rackmount-servers/ and https://eracks.com/tco/
Source: eRacks Open Source Systems
Filed Under: Computers
0 Comments
Latest on The Californer
- Registration Closes October 18 for YMCA Adventure Guides in Simi Valley, Conejo Valley
- International Society of Medical AI Convenes Global Faculty in Florence for ISMAI 2026
- California honored for nation-leading efficiency and data-driven improvements in state services and programs
- Boston Industrial Solutions Introduces Personalized Printing Training
- Mariachi México Lindo Releases Debut Album "Soy de Mexico"
- Century Fasteners de Mexico Exhibiting at the 2026 Automotive & Aerospace Nearshoring Summit
- California: Governor Newsom signs wildfire recovery package to strengthen protections and focus on rebuilding for survivors
- Retro-Bit Announces Legendary Shoot 'em Up R-Type DX as its First Game Boy Color Release
- Stockdale Capital Partners Launches Real Estate Credit Platform
- Retell AI White Label Platform for Agencies Launched by VoiceAIWrapper, With Branded Client Portals and No Per-Minute Markup
- FOCUS Names Mark Phillips Senior Vice President of Business Development
- xBxBio Expands Executive Leadership Search for Next Phase of Cardiovascular Intelligence
- 18th annual Hola México Film Festival presents its Closing Night Film
- Geyser Data Joins NVIDIA Inception as It Builds Cold Data Infrastructure for AI
- Notaron Expands Online Notarization Access Following Wisconsin Approval
- Wings Air Helicopters Selected to Support VIP Transportation for Resorts World in New York
- Tired of Guessing About Your Health? Brittany Poliska Launches 21-Day Wellness Assessment
- Ayurveda, Ayurvedic Medicine & Ayurvedic Science, Dr. Abhay Kumar Pati, USA
- Ayurvedic Medicine, is not only medicine, it is preventative and Curatative, Dr.Abhay K Pati USA
- Scoop Social Co. Brings Its Signature Mobile Dessert Experience to Houston This October