news

WEKA Debuts NeuralMesh 6 to Power Enterprise and Agentic AI Workloads at Production Scale

Jul 22, 2026 · Source: cision
WEKA Debuts NeuralMesh 6 to Power Enterprise and Agentic AI Workloads at Production Scale

The most significant software release in the company's history extends what's possible for agentic AI and other accelerated compute workloads. NeuralMesh 6 redefines inference economics with native multi-tenancy, a unified file-and-object protocol stack built on NVMe, intelligent data mobility with replication and remote caching support, always-on data reduction with performance guarantees, Kubernetes-native operations, and unified observability.

CAMPBELL, Calif., July 21, 2026 -- WEKA, the AI data and memory infrastructure company, today announced WEKA NeuralMesh 6, the most significant software release in the company's history, delivering a purpose-built platform designed to help customers run production AI training, inference, and accelerated compute workloads at production scale on a single unified software stack. The release offers a robust set of capabilities that the AI infrastructure market has historically forced operators to assemble from separate vendors: native multi-tenancy at hyperscale, a full S3 protocol stack on NVMe, intelligent metadata-first data mobility that delivers async replication and remote caching, always-on data reduction with contractual guarantees, Kubernetes-native operations, and unified observability across every deployment.

WEKA NeuralMesh 6 delivers a purpose-built platform to help customers run production AI and accelerated compute workloads.

Inference Costs Now Decide Which AI Companies Scale

The release arrives as the AI industry increasingly shifts decisively from training to production inference. Long-context reasoning, agentic workflows, and retrieval-driven AI workloads place sustained pressure on memory, metadata, and storage in ways traditional infrastructure architectures were not designed to handle. NeuralMesh 6 was built to deliver breakthrough inference economics at production scale for AI clouds, frontier model providers, government agencies, and large enterprises at the forefront of AI innovation.

Inference Economics Proven in Production on Oracle Cloud Infrastructure

NeuralMesh powers hyperscale AI inference in production today on Oracle Cloud Infrastructure (OCI). Leveraging WEKA's Augmented Memory Grid capability, which extends GPU memory by accelerating persistent KV cache access to NeuralMesh-managed NVMe storage, benchmarks on OCI H100 infrastructure have demonstrated 10x higher token throughput, 10x more concurrent users served, and 7x more tokens per GPU in production deployments.

"As agentic AI workloads push context windows and GPU utilization to new limits, WEKA and Oracle Cloud Infrastructure are helping customers scale inference more efficiently without simply adding more GPUs," Pablo Selem, senior director, software development, Oracle Cloud Infrastructure. "WEKA's NeuralMesh platform with Augmented Memory Grid on OCI helps remove memory bottlenecks, delivering substantially more throughput and concurrent users from the same GPU footprint. For customers, that means higher ROI on infrastructure investments and a clearer path to cost-efficient AI at scale."

Inside NeuralMesh 6

NeuralMesh 6 builds on this production foundation. Key features and capabilities include:

Native Multi-Tenancy at Hyperscale

NeuralMesh 6 delivers the industry's only multi-tenancy architecture that combines physical hardware isolation with logical network isolation on a single platform.

  • Composable Clusters provide complete hardware-level isolation, dedicated CPU, memory, and storage drives per tenant for anchor tenants that require guaranteed resources, workload separation, and predictable performance under load.
  • Virtual Multi-Tenancy provides VPC-like network isolation through WEKA's Virtualized RDMA Data Fabric (VRDF), supporting private VLANs, overlapping IP address spaces, per-tenant quality of service (QoS), per-tenant encryption with independent KMS, and independent LDAP/AD authentication. Virtual MT scales to more than 1,000 isolated logical tenants per cluster, with new tenant provisioning in under 30 minutes.

The two tiers compose. A single WEKA hardware cluster running 50 Composable Clusters can support up to 50,000 logically isolated tenants on the same physical infrastructure — a path from dozens of tenants to tens of thousands without re-architecting.

Full S3 Protocol Stack on NVMe

NeuralMesh 6 delivers a fully-featured, native S3 implementation in which the same physical data blocks are addressable via S3 and POSIX simultaneously — not a gateway that translates between protocols, but a single unified namespace. A file written via NFS or POSIX is immediately readable via S3, and vice versa, eliminating the multiple full-dataset copies that a conventional AI pipeline carries between the training, fine-tuning, and inference stages. The implementation is built for AI workload patterns: 2,000 to 5,000 concurrent S3 connections per node, roughly five times the concurrency of conventional S3 architectures. S3 over RDMA enables zero-copy data transfer directly to GPU memory. Object storage is no longer a secondary tier — it is a high-performance protocol native to the same platform that serves POSIX file workloads.

Intelligent Replication for AI Workloads

GPU infrastructure is increasingly distributed across sites, clouds, and regions. Traditional replication architectures force organizations to wait for complete data copies before workloads can begin — a model that breaks down when GPU availability shifts in hours rather than weeks. NeuralMesh 6's intelligent replication changes the model. The same foundation supports AI data mobility, cloudbursting, and multi-site collaboration. Metadata-first replication makes destination environments immediately browsable. Data hydrates on demand, only when accessed, reducing WAN traffic and eliminating unnecessary full-dataset copies. Organizations can place workloads where GPU capacity exists today rather than where data was originally written. In this first step towards complete federation and a global namespace, NeuralMesh 6 unlocks async replication and instantaneous remote caching functionality.

Always-On Data Reduction with Performance Guarantees

NeuralMesh 6 introduces always-on data reduction, specifically designed for AI workloads, with fingerprinting, similarity hashing, deduplication, and compression. It delivers less than 5% write overhead, up to 6x capacity savings on AI training data, and a contractual guarantee of both the data reduction ratio and the performance impact. The performance tax that has historically forced operators to choose between efficiency and speed is gone: data reduction is on by default, every deployment, every workload. Customer-specific pre-sales analysis can project workload-based reduction outcomes before deployment.

AlloyFlash Automated TLC + QLC Flash Tiering

AlloyFlash enables combining TLC and QLC NVMe flash within a single cluster, transparently routing latency-sensitive operations to TLC while leveraging QLC at roughly 30-40% lower cost per terabyte for bulk-capacity workloads. AlloyFlash makes the highest-density WEKApod  configurations production-viable, not just capacity-on-paper.

Automatic, transparent, no customer configuration required.

NeuralMesh Kubernetes Operator

NeuralMesh Kubernetes Operator automates cluster deployment and lifecycle management in Kubernetes environments. For neo-clouds and AI research labs running Kubernetes as their operational standard, Kubernetes Operator reduces deployment time from weeks to hours and brings NeuralMesh into the same declarative operational model their compute infrastructure already uses.

NeuralMesh Observe

NeuralMesh Observe is SaaS-based observability built for high-performance AI storage. Unified multi-cluster dashboards, client-level diagnostics, intelligent alerting with configurable thresholds, and routing to Slack, PagerDuty, or email with direct links to relevant dashboards. Included with every NeuralMesh deployment at no additional cost.

"WhiteFiber is building a distributed GPU platform across multiple data centers connected by high-speed dark fiber. At our scale, data mobility isn't a nice-to-have; it's foundational," said Sam Tabar, CEO at WhiteFiber. "NeuralMesh's intelligent replication makes data mobility real at scale: we can make datasets visible across sites and pull exactly the data each job needs to the next GPU allocation as it becomes available. That shifts replication from a back-end protection function to a core part of how our distributed AI infrastructure needs to operate, improving workload mobility, capacity efficiency, and the resiliency our customers depend on. WEKA's data and memory infrastructure provides the foundation to scale our footprint without compromise. We're excited to keep building on that together."

"The infrastructure operators running production AI today have been forced to assemble platforms from vendors that were never designed to work together. Separate stacks for file and object, manual data movement between them, multi-tenancy bolted on after the fact. NeuralMesh 6 delivers what they've actually needed all along: a single platform that handles the high-performance file layer and the high-capacity object layer on the same blocks, with native multi-tenancy, intelligent data mobility, and always-on data efficiency built in from the start. This is what production inference infrastructure looks like when it's designed for the workload, not retrofitted for it," said Ajay Singh, Chief Product Officer at WEKA.

Availability

NeuralMesh 6 will be generally available in the second half of 2026. Existing WEKA customers can upgrade to NeuralMesh 6 at no additional cost through standard upgrade channels and should contact their Customer Success representative for more details.

For more information about WEKA's NeuralMesh 6 software release and features, visit: https://www.weka.io/product/neuralmesh/.

NeuralMesh x WEKApod: Built to Run In Perfect Symmetry Together

WEKA also announced the third generation of WEKApod Nitro, WEKApod Prime, and WEKApod Prime Max, the first AI storage appliances built on WEKA-designed and engineered hardware to optimize its NeuralMesh software platform. The systems deliver 1.1 exabytes of effective capacity per rack with multiple patents pending. See today's WEKApod announcement to learn more: https://www.weka.io/news/wekapod-3-densest-ai-storage-memory-system.

About WEKA

WEKA is the AI data and memory infrastructure company transforming the economics of agentic AI. Its NeuralMesh™ platform unifies high-performance data storage with extended GPU memory, giving enterprises, AI cloud providers, and AI builders a single foundation for training, inference, and agentic workloads. With Augmented Memory Grid, NeuralMesh extends GPU memory capacity by 1000x, accelerates time to first token by up to 20x, and delivers 10x more concurrent users from the same GPU footprint, proven in production benchmarks. Trusted by 30% of the Fortune 50, WEKA enables organizations to scale AI faster, optimize GPU utilization, and reduce the cost of every token served. Learn more at www.weka.io or connect with us on LinkedIn and X.

WEKA, The Foundation for Enterprise & Agentic AI Innovation

 

Cision View original content to download multimedia:https://www.prnewswire.com/in/news-releases/weka-debuts-neuralmesh-6-to-power-enterprise-and-agentic-ai-workloads-at-production-scale-302828387.html

Comments (0)

Login to join the conversation

Login / Register

No comments yet. Be the first to comment!

More to Read

TECNO Unveiled as Title Sponsor of The SAFF Championship Bangladesh 2026, Bringing AI Innovation to South Asian Football
news
Jul 25, 2026 1 min

TECNO Unveiled as Title Sponsor of The SAFF Championship Bangladesh 2026, Bringing AI Innovation to South Asian Football

As Official Title Sponsor, TECNO joins hands with SAFF to inspire the next generation through football, innovation, and meaningful fan experiences. DHAKA, Bangladesh , July 25, 2026 -- TECNO, an AI-driven innovative technology brand, officially announced its title sponsorship of the SAFF Championship Bangladesh 2026, South Asia s premier international football tournament, during the tournament s official launch ceremony in Dhaka. Scheduled to take place from 4 17 November 2026, the championship will bring together South Asia s leading national teams, celebrating the region s passion for football while strengthening friendship, sporting excellence, and regional unity.

Tony Jaa Becomes GAC's 30-Millionth Customer - GAC Wins Global Trust with "True Craftsmanship"
news
Jul 25, 2026 1 min

Tony Jaa Becomes GAC's 30-Millionth Customer - GAC Wins Global Trust with "True Craftsmanship"

GUANGZHOU, China , July 25, 2026 -- On July 16, at the roll-off ceremony for GAC s 30-millionth vehicle, Feng Xingya, Chairman of GAC Group, handed over the key to the right-hand-drive GAC M8 PHEV (named GN8 overseas) to Tony Jaa. The milestone vehicle is headed straight for overseas markets. Thai action superstar Tony Jaa s choice reflects the trust of 30 million customers worldwide. That trust is built not on showmanship, but on GAC s solid manufacturing true craftsmanship.

From China Mobile's Call Upgrade to the Commercial Launch of "Calling + AI" by Leading Operators: AI Is Reshaping the Value of Native Calling
news
Jul 25, 2026 1 min

From China Mobile's Call Upgrade to the Commercial Launch of "Calling + AI" by Leading Operators: AI Is Reshaping the Value of Native Calling

BEIJING , July 25, 2026 -- On June 15, 2026, China Mobile announced a comprehensive upgrade to its traditional calling services, ushering in a next-generation calling experience defined by HD, intelligence, and security. This milestone not only marks a major leap in telecommunication innovation but also reflects a global, inevitable shift: the transformation of basic communication into intelligent, inclusive services. Breaking Experience Barriers and Redefining the Paradigm of Basic Calling

Reliance Digital Brings Samsung's Latest Galaxy Z Fold8 Series and Galaxy Z Flip8 to Stores Across India
news
Jul 25, 2026 1 min

Reliance Digital Brings Samsung's Latest Galaxy Z Fold8 Series and Galaxy Z Flip8 to Stores Across India

Be among the first to own the new Samsung Galaxy Z Fold8 series and Galaxy Z Flip8. Customers can now pre-order the latest Galaxy foldables at Reliance Digital, with EMIs starting at 6000/month. MUMBAI, India , July 25, 2026 -- Reliance Digital, India s leading consumer electronics retailer, today announced the availability of Samsung s latest generation of foldable smartphones the Galaxy Z Fold8 Ultra , Galaxy Z Fold8 and Galaxy Z Flip8 . Designed to deliver the next evolution of Galaxy AI, powerful performance and iconic foldable innovation, Samsung s newest line-up is now available across Reliance Digital stores and online.

GAC Becomes Official Automotive Partner of Melbourne City FC, Strengthening Commitment to the Australian Market
news
Jul 24, 2026 1 min

GAC Becomes Official Automotive Partner of Melbourne City FC, Strengthening Commitment to the Australian Market

MELBOURNE, Australia , July 24, 2026 -- On 24 July 2026, GAC and Australian professional football club Melbourne City FC officially announced a strategic partnership at AAMI Park. As a key member of City Football Group, Melbourne City FC is one of Australia s leading professional football clubs, with strong competitive performance and an extensive supporter base. Under the partnership, GAC will become the Official Automotive Partner of Melbourne City FC, further strengthening the brand s presence and engagement in the Australian market.

NYSE Content Update: State Street + NYSE to Celebrate 'Fearless Girl'
news
Jul 24, 2026 1 min

NYSE Content Update: State Street + NYSE to Celebrate 'Fearless Girl'

NYSE issues a pre-market daily advisory direct from the trading floor. NEW YORK , July 24, 2026 -- The New York Stock Exchange (NYSE) provides a daily pre-market update directly from the NYSE Trading Floor. Access today s NYSE Pre-market update for market insights before trading begins. Ashley Mastronardi delivers the pre-market update on July 24th Traders continue to monitor escalating tensions in the Middle East. As of 8:00 AM ET, ICE Brent Crude Oil is trading at 98 a barrel. Vertical Aerospace chief engineer David King will join NYSE Live to discuss a flurry of headlines the company made this week. Vertical has joined Project VERTI-GO, designed to accelerate the safe integration of electric aircraft into everyday space. The company also says it completed the first public eVTOL flight in England. The NYSE and State Street will come together to celebrate Fearless Girl this afternoon with a summer block party beginning at 2 PM ET.

STL delivers record financial performance in Q1 FY27; reports highest ever quarterly Revenue, EBITDA and order book
news
Jul 24, 2026 1 min

STL delivers record financial performance in Q1 FY27; reports highest ever quarterly Revenue, EBITDA and order book

Record order book of INR 18,618 Cr EBITDA and PAT hit a record INR 397 Cr and 197 Cr respectively Achieved net debt-free balance sheet MUMBAI, India , July 24, 2026 -- STL (NSE: STLTECH), a leading connectivity solutions provider for AI-ready digital infrastructure, today announced its financial results for the quarter ended 30 June 2026 . For Q1 FY 27 , the Company reported the highest-ever revenues of INR 1,910 Cr and EBITDA of INR 397 Cr , a YoY growth of 87% and 184% , respectively. The open order book stands at INR 18,618 Cr at the end of the Quarter.

HTX Genesis Hackathon Successfully Concludes: Fueling Ecosystem Growth and Long-Term $HTX Value
news
Jul 24, 2026 1 min

HTX Genesis Hackathon Successfully Concludes: Fueling Ecosystem Growth and Long-Term $HTX Value

APIA, Samoa , July 24, 2026 -- After nearly three months of competition, the HTX Genesis Hackathon, co-hosted by HTX DAO and B.AI, with TinTinLand serving as organizer, has successfully concluded. The program progressed from its launch in Hong Kong through project reviews, Demo Day presentations, and the online grand finals, showcasing how global developers and startups are pushing the convergence of artificial intelligence and Web3 beyond theory and into innovative real-world applications and technologies.