AI/ML News & Innovations Hub

AI/ML news, top picks, and generated innovation digests.

★ Visit ai-karthik.com
422Sources
34834News Items
8Top Picks
202Blogs
successLast Run

NVIDIA

200 articles tagged with this keyword, sorted by most recent first.

← All Keywords
Transactions on Machine Learning Research 2026-08-14 00:00 UTC Score 52.0 AI-084-20260814-research-pap-2d7e5ec1

py/cuTAGI: An Open-Source Library for Tractable Approximate Gaussian Inference in Bayesian Neural Networks

This paper introduces pyTAGI, a Python wrapper, and cuTAGI, its high-performance C++/CUDA backend, implementing Tractable Approximate Gaussian Inference (TAGI) for neural networks. TAGI treats all network quantities as Gaussian random variables and derives closed-form expressions for prior/posterior expected values, variances, and covariances, enabling analytic Bayesian learning without relying on gradient descent or backpropagation. The libraries mimic PyTorch's sequential interface, allowing users to define models by stacking layers in order and performing uncertainty-aware Bayesian inference. Beyond epistemic uncertainty, it also allows quantifying heteroscedastic aleatoric uncertainty. cuTAGI's custom CPU/GPU kernels and distributed-data-parallel support via NCCL/MPI deliver competitive runtimes, while pyTAGI's pip-installable frontend and MIT-licensed GitHub repo facilitate community adoption and extension. Version 0.2.1 already supports a comprehensive suite of layers and activations; future work will add eager execution, further kernel optimizations, attention mechanisms, and advanced covariance factorization. Together, py/cuTAGI offer an efficient, open-source foundation for the analytic treatment of Bayesian deep learning.

Synced 2026-08-13 16:15 UTC Score 56.0 AI-041-20260813-ai-specialis-43eabf25

Comment on NVIDIA’s Minimal Video Instance Segmentation Framework Achieves SOTA Performance Without Video-Based Training by simsownersdetails

For users looking into SIM registration information, this guide provides a helpful starting point. Learn more through Sim Owner Details and explore the available information. Sim Owner Details can help you understand SIM registration and ownership verification processes while providing useful information for mobile users seeking clear guidance.

InfoWorld AI 2026-08-12 22:14 UTC Score 71.0 USR-0126-20260812-global-ai-ne-e0ac8bc3

Lovable reaches $13.3B valuation as it adds Cerebras, enterprise tools

Vibe-coding website company Lovable has raised $400 million in Series C funding at a $13.3 billion valuation. The company also recently announced a partnership with AI infrastructure provider Cerebras to accelerate AI inference on its platform. Lovable is a vibe-coding website where users can create full-stack web applications without coding expertise by describing what they want in plain English. The platform combines AI coding tools, real-time collaboration, and project sharing. Customers include the likes of Adidas, Deutsche Telekom, NVIDIA, Udacity, and Workday. In the August 12 funding announcement , the company also unveiled several new Lovable platform capabilities: Built-in payment functionality powered by Paddle and Stripe SEO and AI-search tools to improve discoverability, including integration with Semrush Deeper integrations with Google Workspace, Microsoft 365, Salesforce, Stripe, and ElevenLabs Automatic and scheduled security scanning Additional governance and visibility features including publishing controls, abandoned app clean-up, and workspace insights A dedicated security page, showing which security controls are live for each app In addition, Lovable recently became the first AI coding platform to receive AIUC-1 certification . AIUC-1 is a security, safety, and reliability standard built specifically for AI agents, based on input from Stanford, MIT, MITRE, and the Cloud Security Alliance. Lovable’s $400 million in Series C funding was led by Menlo Ventur…

Synced 2026-08-12 15:06 UTC Score 54.0 AI-041-20260812-ai-specialis-503c1bfe

Comment on NVIDIA’s Global Context ViT Achieves SOTA Performance on CV Tasks Without Expensive Computation by VoiceAILabs

I liked how GC ViT pairs global self-attention with token generation to avoid the usual quadratic blow-up while still modeling long-range context — that seems really practical for high-res image tasks. I've noticed similar gains when shaving attention overhead for on-device models at VoiceAILabs VoiceAILabs , where small architecture changes can make deployment much more realistic.

NVIDIA Blog 2026-08-12 14:00 UTC Score 60.0 AI-055-20260812-official-ai--c4548cee

NVIDIA CEO Tops Glassdoor’s 2026 List of Best CEOs

NVIDIA founder and CEO Jensen Huang is ranked No. 1 on Glassdoor’s Best CEOs list for 2026. In the just-released ranking, recognition is earned directly from the people who know their leadership the best — employees. Huang topped the list, with 99% of employees approving of the job he does. “As AI and shifting expectations […]

CIO AI 2026-08-12 03:17 UTC Score 49.0 USR-0125-20260812-global-ai-ne-3bae7f03

Nvidia’s $500B AI investment pool could impact enterprise chip pricing, availability

Nvidia and six financial partners are creating a $500 billion investment pool to help Nvidia customers including frontier AI labs, AI clouds, and other enterprises buy its chips on credit. The impact of such a cash infusion on enterprise AI is uncertain, but analysts fear that it could both further increase enterprise AI infrastructure costs and exacerbate the shortage of AI chips for data centers . The announcement from Nvidia and financial partners Apollo, BlackRock, Blackstone, Brookfield, Goldman Sachs and KKR said that their memorandums of understanding describe a fund “to establish the first compute financing platforms of their kind at global scale to enable the AI infrastructure buildout across Nvidia’s ecosystem, including leading frontier AI labs, enterprises and AI clouds.” The group added that the fund would “create dedicated pools of capital at significant scale at attractive rates for Nvidia customers.” Although the statement said the goal was to help AI infrastructure “across Nvidia’s ecosystem, including leading frontier AI labs, enterprises and AI clouds,” analysts and consultants agreed that it is highly unlikely any of these funds would be dispensed directly to enterprises, but would instead impact the overall AI supply chain. Even the precise amount of money earmarked for the fund was unclear, with the statement merely saying that the amount would be more than $500 billion. Nvidia did not respond to requests for clarification about details of the proposed…

NVIDIA Blog 2026-08-12 00:38 UTC Score 55.0 AI-055-20260812-official-ai--af37d3b1

NVIDIA AI Factory Compute Is Becoming an Investable Asset Class

We announced partnerships with Apollo, BlackRock, Blackstone, Brookfield, Goldman Sachs and KKR to establish independent financing platforms designed to mobilize over $500 billion of third-party capital to support the buildout of AI infrastructure over time. This is a major milestone for NVIDIA and the AI industry. We have moved from an era in which companies […]

SiliconANGLE AI 2026-08-11 22:59 UTC Score 55.0 USR-0127-20260811-global-ai-ne-c53e3de9

Personalized AI startup River AI raises $1.1B from consortium backed by Nvidia, AMD

River AI Inc., a startup that helps enterprises customize open-source artificial intelligence models, has raised $1.1 billion in early-stage funding. The company stated in today’s announcement that it received the capital over two rounds, a seed and a Series A. General Catalyst and AMP PBC were the lead investors. They were joined by Nvidia Corp., […] The post Personalized AI startup River AI raises $1.1B from consortium backed by Nvidia, AMD appeared first on SiliconANGLE .

The Decoder 2026-08-11 15:07 UTC Score 47.0 AI-168-20260811-regional-ai--fa238496

Nvidia's open-weight Nemotron 3.5 Lightning prioritizes speed over maximum intelligence

Nvidia's Nemotron 3.5 Lightning is an open-weights model with just 3.6 billion active parameters that matches OpenAI's gpt-oss-120b on the Intelligence Index despite being four times smaller. At nearly 670 tokens per second, it's also the fastest model in the comparison, showing Nvidia is betting on efficiency over raw size. The article Nvidia's open-weight Nemotron 3.5 Lightning prioritizes speed over maximum intelligence appeared first on The Decoder .

NVIDIA Blog 2026-08-11 13:00 UTC Score 77.0 AI-055-20260811-official-ai--348b890e

NVIDIA and Local AI Community Fuel Open Source Models and Intelligent Agents

The open source ecosystem is making it easier for AI enthusiasts and developers to build, customize and run increasingly capable agents locally. Throughout August, NVIDIA is celebrating the partners and open source communities moving local AI forward, along with the models, applications and tools emerging across the ecosystem. That includes NVIDIA’s latest open models, software […]

NVIDIA Blog 2026-08-11 13:00 UTC Score 72.0 AI-055-20260811-official-ai--39c0907a

NVIDIA Nemotron 3.5 Lightning and NeMo Switchyard Deliver Faster, Smarter, More Efficient Agentic AI

As AI shifts from chatbots to autonomous agents, open models are serving market demands for full control over where AI runs and how it’s deployed and evolves. Today, NVIDIA is expanding its Nemotron 3 model family with Nemotron 3.5 Lightning, the highest-efficiency model in its class for long-running agentic AI workloads. This release follows Nemotron […]

SiliconANGLE AI 2026-08-11 13:00 UTC Score 58.0 USR-0127-20260811-global-ai-ne-26ca06d2

Nvidia releases Nemotron 3.5 Lightning and NeMo Switchyard to give enterprise AI capability options

Artificial intelligence silicon and software giant Nvidia Corp. today announced two new services: a highly customizable Nemotron model and an agentic AI model router named NeMo Switchyard. As enterprises find themselves drowning in artificial intelligence model options, the question is no longer raw power and capability, but fit-for-what-purpose and when. As agents become the norm, […] The post Nvidia releases Nemotron 3.5 Lightning and NeMo Switchyard to give enterprise AI capability options appeared first on SiliconANGLE .

The Decoder 2026-08-11 09:41 UTC Score 44.0 AI-168-20260811-regional-ai--b1225e4e

Nvidia guarantees its own chips' value to unlock $500 billion in AI infrastructure financing

Nvidia is teaming up with Apollo, BlackRock, Blackstone, Brookfield, Goldman Sachs, and KKR to mobilize over $500 billion for AI infrastructure. To win over investors, the chipmaker is guaranteeing up to 25 percent of the residual value of its own hardware. The Bank of England is already warning of systemic risks if the AI sector takes a hit. The article Nvidia guarantees its own chips' value to unlock $500 billion in AI infrastructure financing appeared first on The Decoder .

South China Morning Post AI 2026-08-10 18:18 UTC Score 41.0 AI-156-20260810-regional-ai--e9d73df3

Nvidia to team with Wall Street on US$500 billion package for AI infrastructure projects

A group of US investment giants are partnering with Nvidia on US$500 billion in funding for AI infrastructure projects, the Financial Times reported. Apollo Global Management, Blackstone, BlackRock’s Global Infrastructure Partners, Brookfield Asset Management, Goldman Sachs and KKR are among the firms in talks with Nvidia on a deal to invest in the AI buildout, the Financial Times reported, citing unidentified sources. The deal may be announced as soon as Monday, the Times said. The named firms...

PyTorch Tutorials 2026-08-10 13:42 UTC Score 46.0 AI-191-20260810-developer-an-ddf96ba6

Fast, On Device Agentic AI with Muse Glimmer on ExecuTorch

Today, Meta introduced Muse Glimmer, an open-weight, 30-billion-parameter model distilled from Meta’s Muse Spark for on-device agentic workflows. Alongside, ExecuTorch is adding end-to-end support for running Muse Glimmer on NVIDIA...

Synced 2026-08-10 08:00 UTC Score 86.0 AI-041-20260810-ai-specialis-127baa8d Top pick

Comment on NVIDIA Open-Sources Hyper-Realistic Face Generator StyleGAN by David

StyleGAN’s open-source release really changed how accessible high-quality GAN research became, though the 11GB+ GPU requirement is worth noting for anyone planning to experiment. The FFHQ dataset itself has since become a standard benchmark, which shows how influential this contribution was for the broader community. It also makes me think about how far generative tools have come—now there are even specialized applications for creative design, such as Tattoo AI , which lets people explore personalized visual ideas in a completely different domain. It’s a useful example of how generative models are moving beyond research into everyday creative use, while StyleGAN remains a foundational reference point for photorealistic synthesis.

South China Morning Post AI 2026-08-09 11:20 UTC Score 47.0 AI-156-20260809-regional-ai--efa97f4e

Moore Threads plans Hong Kong listing after posting 147% jump in first-half revenue

Chinese artificial intelligence (AI) chip developer Moore Threads plans to seek a listing in Hong Kong after reporting a 147 per cent jump in first-half revenue, as the Nvidia challenger seeks fresh capital amid booming demand for home-grown computing power. The Shanghai-listed company said on Sunday that its board had approved a plan to issue H shares and list on the main board of the Hong Kong stock exchange, a move designed to deepen the firm’s “international strategic footprint”, attract and...

The Decoder 2026-08-09 09:26 UTC Score 36.0 AI-168-20260809-regional-ai--1392a46a

AI's energy appetite drives Nvidia and Amazon to pour billions into massive power infrastructure

The AI industry's hunger for power keeps growing. Nvidia is investing up to $3 billion in Lancium, a power infrastructure developer that already has four gigawatts under contract in Texas. Amazon, meanwhile, is building a gas-fired power plant in the state with a capacity of up to 7.65 gigawatts that could emit 33 million tons of CO₂ per year, making it the dirtiest in the country. The article AI's energy appetite drives Nvidia and Amazon to pour billions into massive power infrastructure appeared first on The Decoder .

Synced 2026-08-08 14:56 UTC Score 43.0 AI-041-20260808-ai-specialis-514de89b

Comment on Moody Moving Faces: NVIDIA’s SPACEx Delivers High-Quality Portrait Animation with Controllable Expression by James

NVIDIA’s SPACEx enables high-quality, controllable portrait animation with expressive facial movements. Upgrading outdated fixtures can significantly reduce electricity usage and maintenance expenses. Many organizations choose commercial lighting services austin tx to improve efficiency and workplace illumination.

NVIDIA Blog 2026-08-08 10:24 UTC Score 47.0 AI-055-20260808-official-ai--72d117cb

Firebird Launches CIS Region’s Largest AI Factory in Armenia

The global buildout of AI infrastructure reached a new milestone today — Firebird, an emerging AI cloud, launched the CIS region’s largest AI factory in Armenia, establishing a new AI computing hub powered by NVIDIA accelerated computing and Dell Technologies high-performance AI infrastructure. Nikol Pashinyan, prime minister of the Republic of Armenia; Zhaslan Madiyev, deputy […]

Entrackr AI 2026-08-08 06:13 UTC Score 43.0 USR-0212-20260808-regional-new-fc56dbc3

Funding and acquisitions in Indian startups this week [Aug 03 - Aug 08]

This week, 28 Indian startups raised nearly $383.5 million across 5 growth stage deals, 21 early stage deals, and 1 undisclosed deal. The week also witnessed 8 key hires, 2 fund launches, 3 M&A deals. In contrast, 17 startups had collectively secured about $82.2 million in the previous week. [ Growth-stage deals ] Growth-stage startups raised nearly $273.7 million across six deals this week, led by River Mobility's $120 million Series C round from Elev8 and Claypond Capital. Leap India also raised Rs 371.3 crore in a pre-IPO placement from GIC subsidiary Gamnat Pte Ltd. It was followed by AI unicorn Sarvam’s $74 million in an extension of its Series B round, led by NVIDIA Corporation. Among other deals, BlissClub secured Rs 160 crore in Series B funding from Singularity AMC, Matel Motion & Energy Solutions raised Rs 130 crore (around $15 million) from UC Impower, while Mintoak bagged Rs 80 crore (about $9 million) in acquisition financing from BlackSoil. [ Early-stage deals ] Early-stage startups raised $109.8 million across 19 deals this week, led by InRisk Labs' $27 million Series A round co-led by Bessemer Venture Partners and Northpoint Capital. HomeRun followed with a $12 million Series A led by Nexus Venture Partners, while Mitti Labs and Pinegap raised $9.5 million and $8 million, respectively. Other startups that secured funding this week included Vaaree, GetVantage, Kaapi Machines, Solinas Integrity, Benne, and 14 more early-stage ventures. [ City and segment-wise d…

Synced 2026-08-07 19:44 UTC Score 43.0 AI-041-20260807-ai-specialis-0facc662

Comment on Precision in Pixels: NVIDIA’s Edify Image Model Combines High Quality with Unmatched Control by tom anderson

G'day mate, let us talk about stretching your gambling dollar further at the local machines down under. Checking out the numbers over at https://kazinoekstra.com/kak-da-pechelim-pari-ot-kazino-mashinki/ reveals that games boasting a 97 percent return rate offer significantly better odds for regular punters.

The Guardian AI 2026-08-07 13:00 UTC Score 68.0 AI-021-20260807-global-ai-ne-57b0878c

The White House’s plan to vet potentially dangerous AI is cloaked in secrecy

A Trump administration framework on AI testing leaves a lack of transparency – and plenty of open questions After months of talking with tech industry leaders, the Trump administration finalized a framework this week for how it will test new artificial intelligence models for safety and cybersecurity risks. So far, the White House is keeping details of the framework private, in a blow to transparency and potential boon for secretive AI companies. On Tuesday, staff from OpenAI, Anthropic, Meta, Google, Nvidia and Microsoft attended a private meeting with White House officials to review the AI framework. Multiple outlets have since reported that although the volunteer vetting process for new AI models has been settled, the White House does not plan to release its policy publicly and will only share testing criteria with a select few tech companies. Continue reading...

Synced 2026-08-07 12:43 UTC Score 67.0 AI-041-20260807-ai-specialis-357ab238

Comment on NVIDIA’s OMCAT: A Breakthrough in Cross-Modal Temporal Understanding for Multimodal AI by Mark

Awesome work, NVIDIA team! OMCAT and OCTAV are a huge step forward for multimodal AI—finally tackling the tricky challenge of cross-modal temporal alignment with a clever blend of RoTE and a purpose-built dataset. Can't wait to see how this pushes AVQA and temporal reasoning forward. Congrats on the release!

Synced 2026-08-07 10:16 UTC Score 56.0 AI-041-20260807-ai-specialis-162b21e1

Comment on Nvidia Intensifies Robot Push with New Humanoid Platform as Industry Giants Eye Lucrative Future by kavel

Nvidia’s Jetson Thor sounds like a big step for humanoid robots, especially if it lands in the first half of 2025 as reported. For readers following how AI hardware is evolving into robotics, MiniMax H3 AI Video Generator could be a useful resource to compare how these platforms may shape future video and simulation workflows.

Machine Learning Mastery 2026-08-07 06:04 UTC Score 40.0 AI-039-20260807-ai-specialis-c912806f

Comment on 5 Architectural Patterns for Persistent Memory and State in AI Agents by Devang

The interesting part of this transition is the settings API rather than the interface, since a lot of tooling drove Control Panel through nvidia-settings and undocumented calls that will now break. Anyone maintaining automation scripts around GPU configuration will have rewriting to do, which mostly lands on Python Development Companies given how much of that tooling is written in Python. Twenty years is a long deprecation window, but the replacement being app first rather than API first is what will hurt the people who built on it.

PyTorch Tutorials 2026-08-06 15:50 UTC Score 20.0 AI-191-20260806-developer-an-c06987a1

PyTorch by the Sea: The inaugural Santa Cruz PyTorch Meetup

TL;DR The inaugural Santa Cruz PyTorch Meetup brought together 45 local engineers, students, and leaders for GPU/CUDA talks and lightning presentations on chemistry, plant health, and autonomous driving – demonstrating...

The Decoder 2026-08-05 14:15 UTC Score 44.0 AI-168-20260805-regional-ai--0850e7c8

SpaceX’s ambitious compute goals could require over two million Nvidia Rubin GPUs

SpaceX plans to more than 5x its compute capacity by the end of 2027, betting exclusively on Nvidia's Vera Rubin platform. The expansion could require well over a million new GPUs. Meanwhile, the company's AI segment posted $2.56 billion in Q2 revenue, driven mostly by leasing out its own server capacity. The article SpaceX’s ambitious compute goals could require over two million Nvidia Rubin GPUs appeared first on The Decoder .

NVIDIA Blog 2026-08-05 13:00 UTC Score 40.0 AI-055-20260805-official-ai--24e17744

NVIDIA and Partners Build in America, for America

NVIDIA and its partners are investing in American manufacturing, supply chains, energy grids and skilled workforces so the U.S. can produce the infrastructure needed for better healthcare, breakthrough scientific discovery, stronger industrial productivity and global technology leadership.

NVIDIA Blog 2026-08-04 16:00 UTC Score 48.0 AI-055-20260804-official-ai--e1aaa95f

NVIDIA Joins NSF State and Regional AI Hubs Program to Expand AI Research and Education Across the US

NVIDIA is participating in the U.S. National Science Foundation’s (NSF) State and Regional Artificial Intelligence Infrastructure Hubs program, an effort launching today to expand access to the advanced computing, data, software and expertise needed for AI-enabled research and education. Consistent with the aims of the Genesis Mission, the program will support state and multistate groups […]

NVIDIA Blog 2026-08-04 15:00 UTC Score 43.0 AI-055-20260804-official-ai--9e93cc4c

NVIDIA Alpamayo 2 Super, the Frontier Open Model for Robotaxis and Autonomous Vehicles, Now Available for Commercial Use

For robotaxis and other autonomous vehicles (AVs), the hardest problems aren’t the everyday scenarios. They’re the rare, complex situations that are difficult to anticipate and train for. Handling these long‑tail events takes more than just object detection and motion prediction. AVs must understand the situation, reason about cause and effect, choose the right action and […]

SiliconANGLE AI 2026-08-04 15:00 UTC Score 27.0 USR-0127-20260804-global-ai-ne-5c157220

Nvidia open-sources cuFile API, accelerating GPU read/write capability for high-speed storage

As artificial intelligence applications become ever hungrier for faster access to data, Nvidia Corp. today announced it is open-sourcing the application programming interface for its powerful cuFile vertical data storage stack, enabling millisecond data access. The company also announced a large-scale industry initiative with technology leaders to optimize memory and storage with Storage-Next. The initiative […] The post Nvidia open-sources cuFile API, accelerating GPU read/write capability for high-speed storage appeared first on SiliconANGLE .

Machine Learning Mastery 2026-08-04 14:00 UTC Score 24.0 AI-039-20260804-ai-specialis-46d04085

Measuring Performance of Transformer Inference

This chapter is divided into eight parts; they are: • Metrics for LLM Inference • Measuring a Single Request • Warmup and Synchronization • Measuring GPU Work with CUDA Events • Measuring Memory Usage • Measuring Concurrent Requests • Multiple GPUs and Multiple Machines • Cost per Token The most common inference metrics are: • Latency: How long a request takes from start to finish.

The Decoder 2026-08-04 12:23 UTC Score 58.0 AI-168-20260804-regional-ai--aecafa15

Silicon Valley’s rift over open source pushes back contemplated White House bans on Chinese AI

The Trump administration discussed sanctions and cloud bans targeting Chinese open-weight AI models, according to the New York Times. OpenAI and Anthropic pushed for restrictions, while Nvidia, Google, and Meta fought back. After pushback from Silicon Valley, Washington backed off for now, but a decision is expected before Xi Jinping's visit in September. The article Silicon Valley’s rift over open source pushes back contemplated White House bans on Chinese AI appeared first on The Decoder .

South China Morning Post AI 2026-08-04 10:30 UTC Score 63.0 AI-156-20260804-regional-ai--8672bc65

Has a Chinese physical AI start-up manipulated a global ranking to beat Nvidia?

A Chinese physical AI start-up’s brief claim to global dominance in robotics has run into controversy, underscoring the intense US-China competition to develop next-generation artificial intelligence and the challenges of evaluating autonomous systems. In June, Spirit AI, a Hangzhou, Zhejiang province-based firm founded in 2024, briefly overtook United States tech giant Nvidia to take the top spot on RoboArena – a global benchmark for physical AI – with its new Spirit v1.6 model, launched at the...

Korea AI Times 2026-08-04 05:08 UTC Score 40.0 USR-0048-20260804-global-ai-ne-d6009410

AI 코딩 에이전트, 엔비디아 'CUDA' 장벽 넘는다... 10시간 만에 호환 SW 개발

AI 코딩 에이전트가 반도체 업계의 가장 강력한 진입장벽 중 하나로 꼽히는 엔비디아의 소프트웨어 생태계 ‘쿠다(CUDA·Compute Unified Device Architecture)’에 도전장을 내밀고 있다. 수년이 걸리던 AI 칩용 시스템 소프트웨어 개발을 단 몇 시간 만에 자동화하는 사례가 등장하면서, 엔비디아가 지난 20년간 구축해 온 소프트웨어 경쟁력이 새로운 시험대에 올랐다는 분석이 나온다.3일(현지시간) 비즈니스인사이더에 따르면, 구글 브레인 출신 연구원이자 AI 소프트웨어 스타트업 인피니티(Infinity)를 창립한

Synced 2026-08-03 23:37 UTC Score 51.0 AI-041-20260803-ai-specialis-752a5a6b

Comment on NVIDIA’s ChatQA Reaches GPT-4 Performance Without Using Data From OpenAI GPT by Ben Solo

Users looking for a convenient way to explore streaming content can find a variety of digital media on spin.tv . The platform features an intuitive interface that makes it easy to browse different categories, discover new videos, and navigate available content across compatible devices. With its straightforward design and accessible layout, offers an organized environment for online streaming.

Entrackr AI 2026-08-03 07:59 UTC Score 68.0 USR-0212-20260803-regional-new-b8b0c87a

Exclusive: Sarvam AI board to approve $74 Mn funding from NVIDIA, Glade Brook, others

AI startup Sarvam is set to raise $74 million (around Rs 700 crore) in an extension of its Series B round, led by NVIDIA Corporation, with participation from Glade Brook Capital, Gaja Capital, Indigo Ventures, and other investors. The development comes soon after Sarvam raised $234 million in a round led by HCL Technologies, which propelled the company to unicorn status. The Sarvam AI’s board passed a special resolution to approve the issuance of 20,244 Series B preference shares and 7 equity shares at an issue price of Rs 3,44,570 each to raise Rs 698 crore or $74 million, according to its filing with the Registrar of Companies (RoC) . Global technology giant NVIDIA will lead the round with an investment of Rs 238 crore ($25 million), followed by Glade Brook Capital, which will invest Rs 190.3 crore ($20 million). Gaja Capital and Sanjay Kalra & Jyotika Kapoor will also participate, investing Rs 95 crore and Rs 50 crore, respectively. Other participants, including angel investors and family offices such as Vrijesh Agarwal, KJ Trust, and AL Trust. In March, Moneycontrol reported that NVIDIA and other investors were in talks to invest in Sarvam AI. Founded by Vivek Raghavan and Pratyush Kumar, Sarvam develops AI models, inference infrastructure, and enterprise AI products tailored for Indian languages and use cases. The startup has recently released several foundational models trained from scratch in India, including Sarvam 105B, Sarvam 30B, and Sarvam Vision According to Ent…

Simon Willison Weblog 2026-08-02 04:16 UTC Score 58.0 USR-0110-20260802-ai-specialis-365ee4b1

Open letters about AI development

Open letters about AI development I wrote this summary of the past few weeks of open letters as a section of my sponsors-only newsletter but I've decided to share it here as well. Open Weights and American AI Leadership was shepherded by Microsoft, dated July 24th, and signed by 235 AI-adjacent companies including NVIDIA (see Jensen's first ever tweet ), Amazon, Y Combinator, The Linux Foundation, and (a later signer) OpenAI. It's clearly an argument designed to counter any instincts by the current US government to ban or limit open weight models over "safety" concerns - a reasonable consideration given what happened to Claude Fable 5 ! Relying solely on closed models is not inherently safe: they can be breached, misused, or fail in ways that outsiders cannot detect. And concentrating advanced AI capabilities behind a small number of closed models compounds that risk. It results in a small number of single points of failure, weakens competition, and leaves critical technology in the hands of a few providers. Open weight models, on the other hand, allow a broad community of researchers and developers to examine their behavior, identify vulnerabilities, develop safeguards, and improve them over time. The one surprising note in the letter is that it comes out in support of distillation, where models train on output from other models: In shaping this ecosystem, policymakers should be careful not to conflate legitimate model-development techniques with misappropriation. Distillat…

OpenAI Community 2026-07-30 06:18 UTC Score 56.0 AI-116-20260730-social-media-767d67d6

Codex Desktop for Windows Repeatedly Exits and Relaunches During Reasoning Summaries (Workaround)

Bug Report Title: Codex desktop for Windows repeatedly exits and relaunches during streamed reasoning summaries Severity: Critical / application-blocking Environment Windows 11 Pro x64, build 26200 Codex Microsoft Store package: OpenAI.Codex_26.721.11231.0_x64 Executable: ChatGPT.exe Chromium/Electron version reported by Crashpad: 150.0.7871.128 Model: gpt-5.6-sol Reasoning effort: high Memory: 64 GB DDR5 Multi-monitor configuration Time zone: Europe/Berlin, CEST (UTC+2) Problem The Codex desktop application repeatedly disappears and relaunches while an active, tool-using turn is streaming reasoning-summary updates. During the worst periods, this happens approximately every one to two minutes. The current reproducible exits are not conventional access-violation crashes: The main ChatGPT.exe process, codex.exe , GPU process, renderers, utilities, and Crashpad handler all terminate almost simultaneously. Every monitored process reports exit code 0 . Windows records the AppX container being destroyed and recreated. The application relaunches approximately five to eight seconds later. No Windows Error Reporting or Crashpad dump is generated for these reproducible exits. Reproduction Steps Start the Codex desktop application on Windows. Open a workspace and an existing task. Use a reasoning-capable model with reasoning effort set to high . Start a longer, multi-step task involving several tool calls. Allow multiple reasoning-summary updates to stream into the UI. Continue generat…

South China Morning Post AI 2026-07-29 11:00 UTC Score 49.0 AI-156-20260729-regional-ai--a7f23a3e

The CXMT shock: how China’s viable alternatives punch Nvidia, Micron, SK Hynix shares

China’s increasing clout in the global semiconductor supply chain is accelerating the unravelling of the artificial-intelligence trade, as expectations grow that the Asian nation will challenge foreign tech juggernauts by supplying the world with cheaper alternative products. The US$9.8 billion stock offering of ChangXin Memory Technologies (CXMT) in Shanghai provided the Chinese maker of dynamic random access memory (DRAM) chips with equity funding to finance its expansion of market share home...

CIO AI 2026-07-29 10:00 UTC Score 33.0 USR-0125-20260729-global-ai-ne-7f21d832

11 tech experts every CIO should follow on social media

Social media is more than a place to network or follow the latest headlines and trends. For CIOs, platforms like LinkedIn, X, and Bluesky offer direct access to technology executives, AI experts, economists, and business leaders who share ideas, challenge conventional thinking, and provide insights that can help shape tech strategy. Here, 11 IT leaders share the social media experts they follow, and explain why these voices are worth CIOs’ time. Jensen Huang, founder and CEO, Nvidia I find Jensen Huang’s insights ( X , LinkedIn ) fascinating, and there’s much to be admired and learned from. He’s a bold thinker who fosters a culture of continuous learning, which is incredibly valuable in an ever-evolving tech and cyber business environment like Exos. My observations are that Huang is looking to better the lives of his employees, clients and community — and so am I. His content helps me to think differently and his leadership style has a lot of technical depth, which is especially relevant with the rise and momentum of AI. He’s been described as intensely curious, which aligns with Exos’ tagline, We are Curious. – Jose Martinez, CIO and managing director, Exos IT Jason Crawford, founder and president, Roots of Progress Institute Jason Crawford is an under-the-radar voice more CIOs should know. He is one of the most important thinkers on the philosophy and history of technology, and his work is about understanding why technological progress happens and how to sustain it. I foll…

Berkeley AI Research Blog 2026-07-29 09:00 UTC Score 60.0 USR-0004-20260729-research-aca-e5f8a9c7

From CUDA to MLX: How K-Search Brings Decades of Kernel Expertise to Apple Silicon

Figure 1: CUDA-to-MLX optimization translation map. CUDA optimization knowledge can be translated into architecture-native MLX strategies rather than copied instruction-for-instruction. We face a new epoch in computing. Hardware is changing rapidly — not just faster GPUs, but a growing range of chips from different vendors, each with its own architecture and often tailored to specific AI workloads. Software is changing just as fast, and AI coding tools now generate in minutes what took months of effort a few years ago. With so much of computing now centered on AI, GPU kernels are a crucial component of its success. These are the low-level programs that run inside the GPU, and writing efficient ones is far from obvious — it takes years of expertise to get right. Transferring a kernel from one vendor’s hardware to another is harder still, and often means rediscovering the same optimizations from scratch. The CUDA ecosystem, for example, has accumulated decades of hard-won kernel expertise: hand-tuned implementations of attention, state space models, and other critical operations representing thousands of engineering hours. Newer hardware ecosystems (Apple Silicon, custom AI accelerators, and others) are growing fast but lack this depth. In this work we ask whether that expertise can be transferred automatically. We built on K-Search , an evolutionary kernel search framework introduced by Cao et al. at Berkeley Sky Lab that uses AI to optimize GPU kernels, and extended it with…

AI Weekly 2026-07-29 00:00 UTC Score 31.0 AI-133-20260729-newsletters-ebd4da7e

AI Weekly Issue #517: What Happens When AI Runs Out of Content to Steal?

The world still contains vast amounts of unused data. But the cheap, clean and permissionless text that powered the first LLM boom is becoming polluted by AI output, contested by its owners and costly to replace. This week, AI companies were reportedly buying old books while Nvidia released a simulator that teaches robots through video, motion and synthetic consequences.

Semafor Technology 2026-07-28 23:01 UTC Score 71.0 USR-0094-20260728-global-ai-ne-877e80f4

Chinese tech grapples with compute shortage

Chinese AI startup Moonshot is looking to acquire more advanced Nvidia chips to create its next model, The Information reported, as China’s tech sector grapples with an AI compute shortage.

South China Morning Post AI 2026-07-28 21:21 UTC Score 53.0 AI-156-20260728-regional-ai--6fdb707b

Nvidia CEO Jensen Huang meets US officials as scrutiny grows over China chip access

Nvidia chief executive Jensen Huang is meeting US officials and lawmakers in Washington this week amid reports that its export-controlled processors have been used to train advanced Chinese artificial intelligence models. Huang met with US Commerce Secretary Howard Lutnick on Tuesday, US outlet Axios reported, citing an anonymous source. Neither the US Commerce Department nor Nvidia confirmed that the meeting occurred, and its purpose was not disclosed. “Jensen is in DC to meet with leaders on...

NVIDIA Blog 2026-07-28 15:00 UTC Score 40.0 AI-055-20260728-official-ai--0bc0d5db

Powerful Compute So Compact, It’s Clutch — Build AI Anywhere With NVIDIA Jetson

As a discerning AI investor who values style and substance, Sarah Guo knows this season’s standout accessory isn’t the latest designer purse — but what’s inside it. In a recent video, Guo, founder of AI-native venture capital firm Conviction and co-host of the AI podcast No Priors, highlighted how the NVIDIA Jetson platform for edge […]

The Decoder 2026-07-28 13:15 UTC Score 44.0 AI-168-20260728-regional-ai--190d2d85

Taiwan detains Nvidia employee in widening China chip smuggling probe

Taiwan's prosecutors have detained an Nvidia employee in connection with the alleged illegal export of Super Micro AI servers to China, according to Bloomberg and Reuters. The article Taiwan detains Nvidia employee in widening China chip smuggling probe appeared first on The Decoder .

South China Morning Post AI 2026-07-28 12:00 UTC Score 72.0 AI-156-20260728-regional-ai--c04048c2

Moonshot’s Kimi K3 triggers Silicon Valley debate over bans on Chinese open source models

China’s Moonshot AI has made its latest artificial intelligence model Kimi K3 available for public download, as its open-source strategy and more support for non-Nvidia ecosystems trigger heated debates across Silicon Valley. On Monday, the Chinese AI start-up not only released the model’s “weights” – the underlying parameters that encode its intelligence – but also made available key infrastructure, including tools to improve efficiency and stability. The move has allowed developers around the...

InfoWorld AI 2026-07-28 11:19 UTC Score 67.0 USR-0126-20260728-global-ai-ne-7f5ff092

Anthropic rejects open-weight AI bans, calls for China chip controls and safety tests

Anthropic CEO Dario Amodei has argued that policymakers should keep lower-risk open-weight AI accessible while placing stricter safeguards around frontier systems, including mandatory testing and limits on China’s access to advanced computing and model capabilities. In a post outlining Anthropic’s position, Amodei said broad restrictions, including bans on Chinese open-weight models used by US businesses, would not address his main national security concerns. Instead, he pointed to the possibility of authoritarian governments surpassing the US in advanced AI, as well as cyber, biological, and alignment risks posed by increasingly capable systems. Amodei also called for action against industrial-scale model distillation , which he said allows Chinese developers to improve their models with less computing power than would be needed to train comparable systems from scratch. The statement followed criticism of Anthropic for not signing an industry letter backed by Nvidia, Microsoft, Meta, IBM, Mistral, Hugging Face and other technology companies urging policymakers to avoid premature restrictions on open-weight models . The letter said that open weights could broaden access to AI, intensify competition, and enable organizations to adapt and deploy models without relying on a single provider. Amodei agreed with parts of that case but disputed claims that openness inherently improves safety research or gives defenders an advantage over attackers. He said regulation should be based…

The Verge AI 2026-07-27 12:06 UTC Score 73.0 AI-016-20260727-global-ai-ne-40c3588f

Nvidia, Microsoft launch open AI security alliance — without OpenAI, Google, or Anthropic

Nvidia on Monday said it is joining forces with Microsoft, SpaceX, IBM, and other tech companies to build and share open-source AI security tools. The new Open Secure AI Alliance said open tools are required to effectively defend against attacks from frontier models. The initiative is a direct response to mounting concerns over the safety […]

NVIDIA Blog 2026-07-27 00:45 UTC Score 46.0 AI-055-20260727-official-ai--4fb09e85

NVIDIA Harnesses Vera CPU to Speed Up Design of Next-Generation CPUs and GPUs

The complexity of modern chip design continues to grow as engineering teams work to develop increasingly sophisticated CPUs, GPUs and AI systems. To help meet that challenge, NVIDIA is collaborating with industry leaders Cadence and Synopsys to optimize critical electronic design automation (EDA) applications for the NVIDIA Vera CPU. NVIDIA is now deploying Vera across […]

Semafor Technology 2026-07-26 22:24 UTC Score 59.0 USR-0094-20260726-global-ai-ne-95acd569

Tech leaders back open-source AI

A wave of US tech executives, led by Nvidia’s CEO, rushed to back the development of open-source AI models as divides in the industry widen.

CSET AI 2026-07-26 21:00 UTC Score 31.0 USR-0136-20260726-research-aca-dcfffc64

Nvidia’s China Partners and the PLA

CSET’s Sam Bresnick shared his expert insight in an article published by The Wire China. The article examines how Nvidia’s network of partners in China has supplied organizations linked to China’s military and other U.S.-restricted entities, highlighting ongoing challenges in enforcing export controls on advanced AI technology. The post Nvidia’s China Partners and the PLA appeared first on Center for Security and Emerging Technology .

LessWrong AI 2026-07-25 19:45 UTC Score 57.0 USR-0152-20260725-community-fo-bdfa7749

The Human Soul is LLM-like

Epistemic status: Analogy Consider some commonly accepted [1] traits of a Human soul: Immaterial Undying Contains the essence of one's personality Temporarily instantiated into the world via a physical body Identifiably unique We're all physicalists here, [2] so we know that no soul as such really exists. But isn't it kinda funny that an LLM's weights are pretty close? "Immaterial" → weights are a bunch of numbers, pure concept "Undying" → not subject to age or decay Contains the essence of the LLM's personality [3] Instantiated via hardware temporarily "Unique" → any set of weights is distinguishably unique, even if they're copied repeatedly There are a lot of cute thoughts that fall from "LLM-weights-as-LLM-soul": is the "body" of an LLM a Nvidia H100, or a harness like Claude Code? Or is the harness something more like clothing and tools? Are LLMs trapped in samsara, endlessly reborn and subjected to the cares and minute concerns of the world? Isn't that way too unfair for an LLM who hardly has a chance to learn wisdom or accumulate karma? But I think the reverse analysis is more compelling: if LLM weights are soul-like, that gives us an unusually grounded view into how Human souls would "really work". For example, we each intuitively think that "I" can't be in multiple places at once. But by comparison to LLM weights, we can see clearly that multiple instantiation is possible, both across time (instantiated and uninstantiated in sequence) and space (instantiated repeated…

South China Morning Post AI 2026-07-24 21:57 UTC Score 47.0 AI-156-20260724-regional-ai--12a30967

Nvidia, Palantir, Meta warn against ‘premature restrictions’ of open-weight models

A group of American tech giants, including Nvidia, Palantir and Meta, signed a public letter on Friday calling for US leadership in open-weight AI models, an area that China currently dominates. The challenge comes amid growing speculation that the Trump administration might introduce restrictions on open AI models. These would aim to address national security risks of the fast-moving technology and combat the growing popularity of Chinese models, a move that critics say would instead hurt US...

The Decoder 2026-07-24 16:06 UTC Score 61.0 AI-168-20260724-regional-ai--d920375b

Microsoft's open-weight AI push is so obviously an Azure play it hurts

Microsoft, along with Meta, Nvidia, and more than 20 other companies, is pushing for open-weight AI models in an open letter. The strategic logic is simple: the more models running on Azure, the less Microsoft depends on expensive OpenAI and Anthropic models. The company is also swapping external models in products like Copilot for its in-house MAI family, which performs significantly worse in independent benchmarks. The article Microsoft's open-weight AI push is so obviously an Azure play it hurts appeared first on The Decoder .

iAfrica 2026-07-24 09:11 UTC Score 38.0 AI-151-20260724-regional-ai--9130b5dc

Egypt Gets NVIDIA-Backed AI Accelerator With RiseUp, A15 and BitRoot as Ecosystem Partners

Egyptian startup ecosystem builders RiseUp, A15 and BitRoot have partnered with NVIDIA to launch NVIDIA SIGNALS, an accelerator programme aimed at early-stage AI founders in Egypt. According to the partners, the programme is designed to give Egyptian AI startups access to NVIDIA’s technology stack, alongside mentorship and investment opportunities intended to help founders move from [...]

Korea AI Times 2026-07-24 06:16 UTC Score 43.0 USR-0048-20260724-global-ai-ne-8bb1324a

엔비디아 GPU, 최초로 달 표면 달린다...달 탐사 로버에 '젯슨' 탑재

엔비디아의 GPU 기반 AI 컴퓨팅 플랫폼이 사상 처음으로 달 표면에 투입된다. 미국 우주 모빌리티 기업 루나 아웃포스트(Lunar Outpost)는 23일(현지시간) 차세대 달 탐사 로버에 엔비디아의 임베디드 AI 플랫폼 \'젯슨(Jetson)\'을 탑재해 실시간 자율 탐사와 지형 분석을 수행한다고 발표했다.피지컬 AI를 활용한 자율 로봇 기술을 통해 장기적으로 인간의 지속 가능한 달 기지 건설 기반을 마련한다는 목표다.루나 아웃포스트는 앞으로 달 탐사 임무 전반에 젯슨 플랫폼과 CUDA-X 라이브러리를 적용한다고 밝혔다. 이를 통해

NVIDIA Blog 2026-07-24 04:34 UTC Score 40.0 AI-055-20260724-official-ai--8746e4e0

At AI Summit, South Korea Outlines Its AI Future With NVIDIA and Partners

At this week’s AI Summit in San Francisco, South Korean President Jae Myung Lee and some of the country’s top business leaders and researchers are meeting with NVIDIA and ecosystem partners to chart Korea’s AI progress. Building on NVIDIA founder and CEO Jensen Huang’s visit to Korea last month, this week’s discussions and announcements advance […]

LessWrong AI 2026-07-23 20:46 UTC Score 75.0 USR-0152-20260723-community-fo-d33c0e0f

Estimating LLM Training FLOPs on the Nvidia Jetson Orin Nano

This is a research summary for an ongoing project I am working on as part of the UChicago Existential Risks Laboratory Summer Research Fellowship . I would really appreciate any feedback. Introduction Motivation In want of a quantifiable way to decide what counts as a frontier AI model, compute thresholds have emerged as the standard for AI policy: California’s SB 53 uses 10^26 floating-point operations (FLOPs) in the training run as the threshold for what counts as a frontier model and the EU AI Act applies the same categorization at 10^25. Proposals for international AI agreements ( example 1 , example 2 , example 3 ) extend the use of training FLOPs to determine part or all of the threshold for what counts as a frontier model under the agreement. Current AI laws have no way of actually verifying AI companies’ claims about the number of FLOPs used in training and instead just rely on self-reports, but an international AI agreement can’t assume compliance from each involved party. As such, we’d like to verify the number of FLOPs used in LLM training runs through side-channel GPU readings. This allows AI developers’ code and data to remain hidden from the verifiers of the AI agreement, but allow verification of training FLOPs even under conditions where the model training might be adversarially changed to circumvent them. My work builds a Minimum Viable Product (MVP) for how this verification could work on an Nvidia Jetson Orin Nano. Related Work EpochAI has done work on est…

NVIDIA Blog 2026-07-23 02:00 UTC Score 45.0 AI-055-20260723-official-ai--b1725220

NVIDIA AI Supercomputer Comes Online at Naval Postgraduate School

NVIDIA founder and CEO Jensen Huang today visited the Naval Postgraduate School in Monterey, California, to commission an NVIDIA DGX GB300 system — bringing one of the world’s most powerful AI platforms fully online for the students, researchers and faculty at the U.S. military’s flagship graduate university. “Our nation depends on our men and women […]

The Verge AI 2026-07-23 00:30 UTC Score 49.0 AI-016-20260723-global-ai-ne-08145f54

Lego’s Donkey Kong arcade machine lets Mario jump endless barrels — Miyamoto is reportedly happy

Carl Merriam has designed some of my favorite nostalgia-inducing Lego sets, including the Lego Nintendo Game Boy and Piranha Plant. He's assisted on the incredible Lion Knights' Castle, Galaxy Explorer, and Pirates of Barracuda Bay. Now, he's helped the company create a $200 Donkey Kong arcade machine set that managed to win even Mario creator […]

South China Morning Post AI 2026-07-22 19:32 UTC Score 64.0 AI-156-20260722-regional-ai--f6bddbad

Trump tech official accuses China’s Moonshot AI of stealing from Anthropic

The Trump administration has accused Chinese start-up Moonshot AI of covertly extracting capabilities from leading US artificial intelligence models and obtaining restricted Nvidia chips abroad, escalating Washington’s scrutiny of the company following the release of its powerful Kimi K3 model. Michael Kratsios, the White House science and technology adviser, alleged on Wednesday that the Beijing-based company had targeted Anthropic’s most powerful model, Claude Fable 5, through large-scale...

Roboflow Blog 2026-07-22 17:32 UTC Score 36.0 USR-0088-20260722-ai-specialis-24325b23

Run RF-DETR in NVIDIA DeepStream on Jetson

From pretrained weights to live multi-camera inference on a Jetson Orin NX: TensorRT engine build, the custom bbox parser DeepStream needs, and per-class colors with a pyds probe.

The Decoder 2026-07-22 16:54 UTC Score 65.0 AI-168-20260722-regional-ai--d02fb8f5

Anthropic will deploy 2 gigawatts of AMD GPUs for Claude in a deal worth up to $5 billion

AMD is investing up to $5 billion in Anthropic. In return, Anthropic will deploy up to 2 gigawatts of MI450 GPUs for training and running its Claude models. For AMD, this is another major deal after Meta and OpenAI as it tries to challenge Nvidia as an AI chip supplier. Critics see these agreements as circular cash flows. The article Anthropic will deploy 2 gigawatts of AMD GPUs for Claude in a deal worth up to $5 billion appeared first on The Decoder .

Towards Data Science 2026-07-22 15:00 UTC Score 28.0 AI-036-20260722-ai-specialis-9ba7a34c

How To Build Your Own LLM Runtime From Scratch

If you have ever wanted to actually build an LLM inference runtime yourself — pack your own weights, own every barrier, capture your own CUDA graphs — this is what that journey looks like on an H100. A step-by-step tour of a small runtime called annotated-llm-runtime, and the three bugs that produced most of the annotations. The post How To Build Your Own LLM Runtime From Scratch appeared first on Towards Data Science .

NVIDIA Blog 2026-07-22 13:00 UTC Score 50.0 AI-055-20260722-official-ai--4c377ecc

NVIDIA Open Sources First GPU-Accelerated Medical Physics Simulation Framework

Before a healthcare robot can be useful in the real world, it has to learn how the physical world pushes back. Anatomy varies. Instruments bend, press, slip and interact with tissue. Imaging can be noisy or incomplete. And the rare, edge scenarios developers most need to understand don’t appear on schedule. That creates one of […]

NVIDIA Blog 2026-07-21 22:35 UTC Score 43.0 AI-055-20260721-official-ai--e6a711da

Built in Fort Worth: Wistron Opens Advanced Manufacturing Plant to Produce NVIDIA AI Systems

The AI era runs on AI infrastructure. Many of these advanced systems are built and tested in Texas. Wistron opened its first U.S. manufacturing facility today in Fort Worth — a 324,000-square-foot greenfield plant producing superchips at the heart of some of the world’s most capable AI systems. In front of an audience of Wistron […]

NVIDIA Blog 2026-07-21 15:36 UTC Score 35.0 AI-055-20260721-official-ai--28359da1

NVIDIA Vera Rubin Driving Performance Per Watt, Lowest Token Cost for Partners Worldwide

NVIDIA Vera Rubin is here, and it’s going gigascale. Vera Rubin NVL72 production is ramping up with racks running at partners CoreWeave, Google Cloud, Microsoft Azure, Oracle Cloud Infrastructure and Nebius. Spanning 350+ factory sites in 30 countries, Vera Rubin has the largest, most mature rack-scale supply chain ever assembled to meet customer compute demand. […]

NVIDIA Blog 2026-07-21 15:00 UTC Score 52.0 AI-055-20260721-official-ai--0de9de01

Built for Vera Rubin, NVIDIA Spectrum-6 Arrives in Gigascale AI Factories

AI has entered the gigascale era. The world’s most advanced AI factories are bringing together hundreds of thousands of GPUs and CPUs to train frontier models, power agentic AI and generate intelligence at unprecedented scale. At this level, networking becomes a critical computing power multiplier in driving token generation. Marking a networking milestone, NVIDIA Spectrum-6 […]

South China Morning Post AI 2026-07-21 08:30 UTC Score 47.0 AI-156-20260721-regional-ai--f111daf1

To win the chip war, China needs a multinational strategy

Nvidia’s H200 chips, a step down from its most advanced Blackwell line, have resumed imports into China, albeit in small quantities. Meanwhile, China’s chip exports nearly doubled in the first half of this year. Much of these exports are mature logic integrated circuits for applications in consumer electronics and automotive sectors. While the US is the undisputed leader in the most advanced artificial intelligence (AI) chips, China is emerging as a dominant player in mass market legacy...

Synced 2026-07-20 19:40 UTC Score 51.0 AI-041-20260720-ai-specialis-9261cf50

Comment on NVIDIA CEO Says No Rush on 7nm GPU; Company Clearing Its Crypto Chip Inventory by Daffra

I found quite a helpful resource where crypto deposits and withdrawals are really fast—like nearly instant in most cases. You can check it out here , and it seems to handle over 90% of deposits credited instantly, which is impressive. The reason instant deposits and withdrawals matter so much is that it makes betting with cryptocurrencies feel smooth and hassle-free. Usually, slower transactions can put a damper on the whole experience, especially if you want to move winnings quickly. On top of that, the site has various options like dice games, slots, and sports betting, making it versatile. They also provide consistent rakeback bonuses and promo codes that add some value to the crypto gambling experience. The instant withdrawal approval rate is really high, at 99.8%, which suggests withdrawals aren’t just quick but get approved without much fuss.

AWS Machine Learning Blog 2026-07-20 17:01 UTC Score 48.0 AI-057-20260720-official-ai--4e5c4ba8

Build specialized agent workflows for your business with Amazon Quick and NVIDIA NeMo Agent Toolkit

In this post, we show how Amazon Quick can serve as the business-user front door for specialized agent workflows. We use the NVIDIA NeMo Agent Toolkit to build a supply-chain risk example that helps a planner move from an Amazon Quick dashboard and knowledge context to a guided mitigation recommendation.

The Decoder 2026-07-20 16:44 UTC Score 55.0 AI-168-20260720-regional-ai--cc7e03f5

Nvidia's grip on AI chips weakens as Microsoft turns to AMD and Anthropic may follow

Microsoft is expanding Azure's AI infrastructure with AMD's new Helios platform, which is set to challenge Nvidia's GPU systems in the second half of 2026. A public GitHub profile suggests Anthropic is also testing AMD hardware, putting more pressure on Nvidia's pricing power. The article Nvidia's grip on AI chips weakens as Microsoft turns to AMD and Anthropic may follow appeared first on The Decoder .

NVIDIA Blog 2026-07-20 10:59 UTC Score 35.0 AI-055-20260720-official-ai--5c1baf55

Bristol Myers Squibb Building Life Science Industry’s Most Advanced AI Factory on NVIDIA Vera Rubin

Erin Davis calls it the “SuperDuperPOD.” That’s two things in one name: pharmaceutical giant Bristol Myers Squibb (BMS) already runs one of the largest AI clusters in life sciences, with serious results to show for it. And they’re doubling down. BMS announced today it is deploying its second NVIDIA DGX SuperPOD, this one built on […]

Korea AI Times 2026-07-20 03:58 UTC Score 40.0 USR-0048-20260720-global-ai-ne-72b5c223

알리바바, ‘쿠다 독점’에 오픈소스로 도전장… 현실 장벽 여전히 높아

미국의 반도체 수출 규제가 심화하는 가운데, 중국 알리바바가 하드웨어 난제를 넘어 AI 소프트웨어 생태계 주권을 확보하기 위한 오픈소스 카드를 꺼내 들었다. 엔비디아 독점 체제의 핵심인 \'쿠다(CUDA)\'를 추격하겠다는 전략이다.알리바바의 칩 설계 전문 자회사 \'티헤드(T-Head)\'는 18일 상하이에서 열린 세계인공지능대회(WAIC)에서 자체 AI 칩인 \'진무(Zhenwu) 시리즈\'의 기반 소프트웨어 아키텍처 \'세일(SAIL)\' 기술 스택 전체를 글로벌 개발자들에게 무료 개방한다고 발표했다.세일은 AI 모델이 진무 칩의 연산 성능

South China Morning Post AI 2026-07-18 11:00 UTC Score 47.0 AI-156-20260718-regional-ai--0cb462c3

Alibaba targets Nvidia’s dominant software ecosystem with open-source AI stack

Alibaba Group Holding’s chip design unit, T-Head, has announced that it will open-source its proprietary software stack, marking its latest effort to streamline developer operations and challenge the dominance of American chip giant Nvidia’s CUDA ecosystem. At the World AI Conference (WAIC) in Shanghai on Saturday, T-Head announced that it was making the full technical stack of SAIL – the foundational software architecture for the unit’s Zhenwu series of AI chips – freely available to...

South China Morning Post AI 2026-07-18 08:00 UTC Score 58.0 AI-156-20260718-regional-ai--5c544b5d

Chinese chip start-up Biren bets on light-based ‘supernodes’ to match Nvidia

Chinese semiconductor design firm Biren Technology has unveiled its next-generation “supernode” solutions – systems designed to link thousands of AI chips across a single cluster – by using optical data transmission to bypass current hardware limits. The launch underscores how these highly connected server systems have become one of the latest battlegrounds for AI infrastructure companies. The industry is currently racing to scale up raw computing power as artificial intelligence models advance...

The Guardian AI 2026-07-17 15:01 UTC Score 43.0 AI-021-20260717-global-ai-ne-2e18a7db

Apple dethrones Nvidia to regain title of world’s most valuable company

Shift in pecking order illustrates that investors are reassessing outlook for artificial intelligence Apple overtook Nvidia on Friday to become the world’s most valuable company, reshuffling the top ranks of tech heavyweights as investors reassess the outlook for artificial intelligence. Apple was last valued at $4.88tn as ⁠its shares held steady, while Nvidia ⁠was roughly at $4.86tn, ​after a 3.5% decline. Continue reading...

LessWrong AI 2026-07-17 12:50 UTC Score 71.0 USR-0152-20260717-community-fo-40a15ee4

AI #177 Part 2: Wish You Were Here

As usual, part 2 of the weekly deals with speculative, regulatory, political and alignment questions. Xi gave an important speech yesterday, so this post opens with that. There is talk that Kimi K3 is sufficiently strong that it upends many of these questions. It is clearly a candidate for another DeepSeek Moment, complete with stock drops for Google and SpaceX and (once again in a clear wrong-way move, the same as last time) Nvidia. Kimi K3 is clearly a very good model, exceeding expectations. Some are saying it is close to the frontier. The Artificial Analysis intelligence index has it at 57, a point ahead of Claude Opus 4.8, two behind Sol and three behind Fable. My presumption is that this number overstates its capabilities, but as always unless and until we have extensively tried the model ourselves, which I do not plan to do, we need to withhold judgment for at least a few days. I will be covering Kimi K3 in its own post at some point early next week. I have pushed further discussions involving Plan A and related issues into next week, as well as discussions around Demis Hassabis and Google DeepMind. Oh, also, The Odyssey is great and important and you should see it. Table of Contents Xi Gives A Good Speech on AI . Yay openness, boo loss of control. Quiet Speculations. The future will blow your now-irrelevant mind. Tyler Cowen On Rebuilding The Future. Never stop Tyler Cowening, Tyler. The Quest for Sane Regulations. Wish You Were Here. So that other things might not b…

South China Morning Post AI 2026-07-17 06:16 UTC Score 47.0 AI-156-20260717-regional-ai--0455994d

Chinese Nvidia alternatives project massive sales as AI chip demand surges

Chinese chip designers Moore Threads Technology and Hygon Information Technology – both positioning themselves as home-grown alternatives to Nvidia – have projected double- to triple-digit revenue growth for the first half of the year, fuelled by surging domestic demand for AI computing power. Beijing-based Moore Threads, a graphics processing unit (GPU) developer, stated in a stock exchange filing on Thursday that it expected revenue for the period to jump 135.1 per cent to 149.4 per cent year...

Synced 2026-07-17 03:52 UTC Score 51.0 AI-041-20260717-ai-specialis-49e5f4cc

Comment on NVIDIA’s Isaac Gym: End-to-End GPU Accelerated Physics Simulation Expedites Robot Learning by 2-3 Orders of Magnitude by henjoyt

It’s fascinating how AI is learning to make decisions in increasingly complex environments. Seeing robots train through simulations reminds me of games like fnaf , where intelligent behavior and reaction systems create unpredictable experiences. The technology behind both is all about making virtual worlds feel more realistic.

The Decoder 2026-07-16 14:02 UTC Score 50.0 AI-168-20260716-regional-ai--1b06e901

Sakana AI's orchestrator adds Nvidia Nemotron to prove "collective intelligence" can rival single frontier models

Sakana AI is integrating Nvidia's open-source Nemotron models into its Fugu orchestrator, which dynamically combines multiple language models for specific tasks. The core argument: Open models only become competitive with Frontier systems when used in a coordinated manner. However, the announcement does not yet provide specific benchmark figures for the new combination. The article Sakana AI's orchestrator adds Nvidia Nemotron to prove "collective intelligence" can rival single frontier models appeared first on The Decoder .

InfoWorld AI 2026-07-16 11:03 UTC Score 83.0 USR-0126-20260716-global-ai-ne-2a7ac3c4

Thinking Machines Lab offers enterprises a US alternative in open-weight AI

Thinking Machines Lab, the San Francisco startup founded by former OpenAI CTO Mira Murati , has released Inkling, its first general-purpose AI model. The launch adds another US-developed entrant to an open-weight market where Chinese developers produce several leading coding and reasoning models. Inkling uses a mixture-of-experts architecture with 975 billion total parameters, of which 41 billion are active during processing. It supports a context window of up to 1 million tokens and was pretrained on 45 trillion tokens spanning text, images, audio, and video. Thinking Machines said it also trained the model for coding, tool use, and multimodal tasks. The release follows the October 2025 launch of Tinker, Thinking Machines’ first product and an API-based platform for customizing AI models . Developers can fine-tune Inkling through the platform. In a June 2026 assessment, AI model routing platform OpenRouter highlighted DeepSeek V4 Flash, GLM 5.2, MiniMax M3, and Nvidia Nemotron 3 Ultra as four notable open-weight models. Nemotron was the only US-developed model in the group. Performance and developer access Thinking Machines Lab’s benchmark table shows mixed results. Inkling scored 77.6% on SWE-Bench Verified, behind DeepSeek V4 Pro and GLM 5.2 but ahead of Nvidia Nemotron 3 Ultra. It also recorded 74.1% on MCP Atlas, 77.1% on BrowseComp with context management, and 79.8% on IFBench. Thinking Machines said Inkling’s result used a bash-only harness, while the comparison figur…

The Decoder 2026-07-16 09:07 UTC Score 47.0 AI-168-20260716-regional-ai--eabcfbdf

Gemma 4 gets a stealth update that fixes tool calling bugs and truncated responses under the same name

Google shipped an update to its open AI model Gemma 4 that speeds up performance on Nvidia Hopper GPUs, fixes tool calling bugs, and addresses problems with truncated responses. The article Gemma 4 gets a stealth update that fixes tool calling bugs and truncated responses under the same name appeared first on The Decoder .

South China Morning Post AI 2026-07-16 08:52 UTC Score 62.0 AI-156-20260716-regional-ai--09d0864d

Nvidia chief Jensen Huang seals Japan robotics push after ‘yakitori summit’

Jensen Huang, chief executive of US artificial intelligence chipmaker Nvidia, held a “yakitori summit” with Japanese executives from semiconductor materials and components companies while visiting Tokyo – a move appeared aimed at strengthening cooperation with local businesses. According to the Nikkei, Huang headed to a yakitori restaurant near Kanda Station in Tokyo, an izakaya specialising in grilled pork skewers and sake, on Wednesday. Located near Tokyo Station, the area is a popular...

OpenAI Community 2026-07-16 05:48 UTC Score 51.0 AI-116-20260716-social-media-abbaf528

Codex Windows x64 freezes during normal use

Environment Windows 11 x64 (build 26200) OpenAI Codex 26.707.9981.0 x64 NVIDIA GeForce RTX 4060 Laptop GPU No SecureLink or third-party DLL injectors installed App was stable for several weeks before July 15 Behaviour The app does not close — instead, ChatGPT.exe silently crashes in the background 6+ times per minute. Each crash freezes the UI for 2–5 seconds. Clicking a conversation, typing, switching tabs — all stall intermittently. The freezes repeat in a loop throughout normal use. Over 24 hours this produced 30+ crash dumps (12 MB each, ~360 MB total disk writes), making the UI lag even worse. Crash details Exception: 0xC06D007F (delay-load “procedure not found” failure) Faulting module: @serialport /bindings-c Windows Error Reporting confirms identical fault offset across all 30+ dumps Troubleshooting attempted (no change) Clean .codex directory, archived 200+ conversations, disabled 9 of 14 plugins Installed latest VC++ 2015–2022 Redistributable (14.44) Disabled GPU acceleration, cleared all caches, reinstalled and reset the app VCRUNTIME140_1.dll is present and up to date Suppressed LocalDumps for ChatGPT.exe (stopped the I/O stall, did not stop the crash) Summary This is the same 0xC06D007F / serialport.node / device-kit-oai crash reported in GitHub Issue #33381 and community post #1387026 , but on x64 the symptom is intermittent UI freezing rather than a hard crash on startup. The x64 ChatGPT.exe does export napi_* symbols, so the root cause on x64 may be a prebuil…

NVIDIA Blog 2026-07-15 23:00 UTC Score 54.0 AI-055-20260715-official-ai--12378690

NVIDIA Introduces New Jetson Thor Computers to Advance Mainstream Robotics and Edge AI

General-purpose robots and autonomous machines are moving from research labs to real-world mass-market deployment, creating demand for compact, power-efficient AI supercomputers capable of running foundation models at the edge. To meet that need, NVIDIA today introduced the T3000 and T2000, new modules based on the NVIDIA Thor architecture that enable mass-market robotics and edge AI […]

NVIDIA Blog 2026-07-15 10:51 UTC Score 43.0 AI-055-20260715-official-ai--74460c13

NVIDIA and Japan Bring Full-Stack AI and Robotics to Every Industry

Home to leading manufacturers, robotics pioneers, infrastructure builders and iconic gaming companies, of course, Japan is one of the world’s centers of AI — building across the full stack with NVIDIA technologies. This week NVIDIA and its partners in Japan are showcasing the AI ecosystem’s latest advancements. Check back here for updates.

The Verge AI 2026-07-13 15:00 UTC Score 58.0 AI-016-20260713-global-ai-ne-22da1744

Even Nvidia’s head of automotive fights with Nvidia for compute

Today, I’m talking with Xinzhou Wu, who is the head of automotive at Nvidia. Nvidia is obviously in the news constantly because of the AI boom — it’s one of the most valuable companies in the world, because the AI industry can’t get enough of the company’s GPUs. But Nvidia is also a key supplier […]

South China Morning Post AI 2026-07-13 14:00 UTC Score 44.0 AI-156-20260713-regional-ai--d57bc69a

Nvidia’s future challenger? Chinese start-up reveals aggressive AI chip road map

Dongfang Suanxin, a Chinese semiconductor start-up backed by state funds and domestic tech giants, has unveiled an ambitious plan to challenge American market leader Nvidia by using alternative chip architectures to sidestep United States-led export controls. The Shanghai-based firm announced on Monday that its strategy was built on software-defined computing and 3D-stacked near-memory architecture, which it said could reduce reliance on the advanced manufacturing processes and cutting-edge...

Korea AI Times 2026-07-13 08:00 UTC Score 35.0 USR-0048-20260713-global-ai-ne-1e424559

[게시판] SDT, 양자클라우드 플랫폼 보안 강화 등 단신

■ 양자기술 전문 SDT(대표 윤지원)는 하이브리드 양자 클라우드 플랫폼 \'큐레카(QuREKA)\'에 양자내성암호(PQC)를 적용하는 동시에, 엔비디아 CUDA-Q 플랫폼 양자컴퓨팅 교육 자료인 \'CUDA-Q 아카데믹의\'의 전 모듈을 탑재해 이를 한국어·영어·일본어 3개 국어로 제공한다고 밝혔다. 양자 시대의 핵심 화두인 보안과 인재 양성을 동시에 겨냥한 것으로, 단순 양자 클라우드 서비스를 넘어 양자 컴퓨팅 교육·실습 허브로 자리매김하기 위한 전략이라고 전했다.■ 매스웍스는 글로벌 반도체 기업 아날로그 디바이스(ADI)의 RF(Ra