Neura News

AI News

News reporting focused on AI and machine learning, covering the companies behind these technologies, their real-world applications, and the ethical concerns they raise. This includes areas like generative AI (large language models, text-to-image and video), speech tech, and predictive analytics.

Latest News

35 articles
Funding

Aranya launches with $11M to turn bare-metal servers into GPU clusters in under 48 hours

Aranya Inc. launched with $11M in funding to convert bare-metal servers into production-ready GPU clusters for AI inference in under 48 hours. Its open-source clusterdOS, built on Kubernetes, automates hardware discovery, diagnosis, and remediation, addressing Kubernetes' blind spot in hardware failures. The company has managed over $500M in GPU hardware and promises a 48-hour deployment timeline, with security and natural-language interface features.

Sep 14 minNeura News
Industry

U.S. Moves to Close Cloud-Compute Loophole That Lets China Rent Banned GPUs

The Trump administration is drafting legislation to close a loophole allowing China to rent advanced GPU computing power via cloud services in third countries like Vietnam and Singapore. The Remote Access Security Act and a House version would extend export controls to remote access, requiring U.S. cloud providers to verify user identities and block entities linked to China's military or AI programs. The move aims to slow China's AI progress through attrition, though enforcement remains challenging.

Aug 296 minNeura News
AI Models

Nvidia's Nemotron 4 Targets Trillion-Parameter Scale, but Chinese Rivals Already Lead

Nvidia is developing Nemotron 4, an open-weight AI model family with a trillion-parameter flagship, expected as early as fall 2026. However, Chinese labs like Moonshot AI and DeepSeek already ship larger models, with Kimi K3 at 2.8 trillion parameters and DeepSeek V4 Pro at 1.6 trillion. Nvidia has tripled cloud spending to $28 billion through 2031, betting on open models to drive GPU demand, even as it competes with customers like OpenAI.

Aug 123 minNeura News
AI Tools

NVIDIA Makes the Case for AI Factories as an Investable Asset Class

NVIDIA argues that AI factory compute should be treated as an investable infrastructure asset class, citing rising GPU rental prices and long hardware lifespans. The company recently announced partnerships with six financial institutions to mobilize over $500 billion in third-party capital for AI infrastructure. The blog post details the economics behind these partnerships, including residual-value support and independent underwriting.

Aug 126 minNeura News
Developer

Meta's Muse Glimmer Launch and Zuckerberg's Personal Superintelligence Essay Mark a Return to Open-Weight Leadership

Meta released Muse Glimmer, an open-weight 30B-parameter multimodal agent model under Apache 2.0, optimized for local deployment on consumer GPUs. CEO Mark Zuckerberg published an essay on Personal Superintelligence, outlining Meta's vision for accessible AI and addressing labor, geopolitics, and compute. The release signals Meta's return to open-weight leadership.

Aug 1113 minNeura News
Industry

Meta's Muse Glimmer Brings Agentic AI to the Desktop, Sparking a Cloud vs. Local Cost Debate

Meta released Muse Glimmer, a 30-billion-parameter AI model designed to run locally on a single GPU, enabling always-on agentic workflows. The release has sparked debate over whether enterprises should shift from cloud-based AI to on-premises hardware, with analysts noting complex cost comparisons between capital expenditure and operational expenses. While local deployment offers control and predictable costs, quantization and hidden hardware costs complicate the financial decision.

Aug 118 minNeura News
AI Tools

NVIDIA Launches Magpie Multilingual TTS, an Open-Weight Voice Model for 12 Languages

NVIDIA has released Magpie Multilingual TTS, an open-weights text-to-speech model supporting 12 languages, including new additions like Arabic, Korean, and Brazilian Portuguese. The 364M-parameter model is designed for low-latency voice agents, offering developers full control over deployment, from private servers to air-gapped environments. Benchmarks show sub-200ms end-to-end latency on B200 GPUs, with architectural improvements like frame stacking and local transformers enhancing speed. The model is available on Hugging Face under the NVIDIA Open Model License, paired with NIM microservices and NeMo for fine-tuning.

Aug 107 minNeura News
Industry

Firebird Opens CIS Region's Largest AI Factory in Armenia

Firebird, a U.S.-based AI cloud company, opened the CIS region's largest AI factory in Hrazdan, Armenia, on August 8, 2026. Built on NVIDIA accelerated computing and Dell PowerEdge servers, the facility is scaling from 15 MW to 300 MW of AI infrastructure capacity, with plans to deploy over 70,000 NVIDIA GPUs by 2027. The opening ceremony drew high-level officials from Armenia, Kazakhstan, and the U.S., highlighting the project's regional significance.

Aug 86 minNeura News