Warp Brings xAI's Grok 4.6 to Its Terminal, Matching GPT-5.6 on Key Benchmarks
Grok 4.6, xAI's frontier-class agentic model, lands in Warp Terminal and the new Warp Agent CLI via X Premium or SuperGrok
AI/ML news, top picks, and generated innovation digests.
157 articles tagged with this keyword, sorted by most recent first.
Grok 4.6, xAI's frontier-class agentic model, lands in Warp Terminal and the new Warp Agent CLI via X Premium or SuperGrok
plus skills and tools to try with agents
PLUS: Pixel 11, Claude Voice, and a $2B enterprise AI bet.
์ผ๋ก ๋จธ์คํฌ CEO๊ฐ ์คํ์ด์คX์ ๋ฐฉ๋ํ ๋ด๋ถ ๋ฐ์ดํฐ๋ฅผ AI ๋ชจ๋ธ โ๊ทธ๋ก(Grok)โ ํ์ต์ ํ์ฉํ๊ฒ ๋ค๋ ๊ณํ์ ๋ฐํ๋ค. ๋น์ฆ๋์ค ์ธ์ฌ์ด๋์ ๋ฐ๋ฅด๋ฉด, ๋จธ์คํฌ CEO๋ 11์ผ(ํ์ง์๊ฐ) ์ด๋ฆฐ ์ ์ฒด ํ์์์ โ์คํ์ด์คX์ ๋ชจ๋ ์ ๋ณด๋ฅผ ํฉ์น ๋ฐ์ดํฐ๋ก ๊ทธ๋ก์ ํ์ต์ํฌ ๊ฒโ์ด๋ผ๋ฉฐ โ์ด๋ค ์๋ฏธ์์๋ ๊ทธ๋ก์ด ์ฌ๋ฌ๋ถ์ ํ์ตํ๊ฒ ๋ ๊ฒโ์ด๋ผ๊ณ ๋งํ๋ค.๊ทธ๋ ์คํ์ด์คX ์ง์๋ค์ โ์ง๊ตฌ์์์ ๊ฐ์ฅ ๋ฐ์ด๋ ์ธ๊ฐ๋ค์ ์งํฉโ์ด๋ผ๊ณ ํํํ๋ฉด์, ์ง์๋ค์ ์๊ฐ๊ณผ ์์ด๋์ด, ๊ฐ์น๊ด์ด ์์ผ๋ก AI ๋ชจ๋ธ์ ๋ฐ์๋ ์ ์๋ค๊ณ ์ค๋ช ํ๋ค.ํนํ ์ง์๋ค์ด AI์ โ๋ถ๋ชจโ ์ญํ
์คํ์ด์คXAI๊ฐ ์ฐจ์ธ๋ AI ์์ด์ ํธ ์์ฅ์ ๊ฒจ๋ฅํ ๊ณ ์ฑ๋ฅ ๋ชจ๋ธ์ ๋ด๋์ผ๋ฉฐ ์คํAI์ ์คํธ๋กํฝ์ ์ถ๊ฒฉํ๊ณ ์๋ค. ํนํ ์ํฐํผ์ ์ ๋๋ฆฌ์์ค์ ์ง๋ฅ ์์(AAII)์์ 3์๊ถ์ ์ง์ , ๋ณธ๊ฒฉ์ ์ธ ํ๋ก ํฐ์ด ๋ชจ๋ธ ๊ฒฝ์์ ํฉ๋ฅํ๋ค. ์คํ์ด์คXAI๊ฐ ์ฅ์๊ฐ ์์ด์ ํธ ์์ ๊ณผ ์ฝ๋ฉ, ์ง์ ์ ๋ฌด์ ํนํํ ์ต์ AI ๋ชจ๋ธ โ๊ทธ๋ก 4.6(Grok 4.6)โ์ ๊ณต๊ฐํ๋ค.๊ทธ๋ก 4.6์ ์ฌ๋ฌ ๋จ๊ณ์ ๊ฑธ์น ์ฅ๊ธฐ ์์ ์ ์ง์์ ์ผ๋ก ์ํํ๋ ๋ฅ๋ ฅ์ ์ด์ ์ ๋ง์ท๋ค. ๋จ์ํ ์ง๋ฌธ์ ๋ตํ๋ ๊ฒ์ ๋์ด ์๋ก์ด ๋ถ์ผ๋ฅผ ์กฐ์ฌํ๊ณ ์ ๋ณด๋ฅผ ๋ถ์ํ๊ฑฐ๋ ์ฝ๋๋ฒ ์ด์ค๋ฅผ ํ์ํ๊ณ , ์ ํ
AI teammate category just had its most significant new entrant yet
SpaceXAI today released Grok 4.6, a large language model that it says can outperform Anthropic PBCโs Claude Fable 5 in some areas. SpaceXAI was known as xAI until last month. The Elon Musk-founded artificial intelligence provider rebranded in connection with its acquisition by SpaceX Corp. In June, the combined company listed its shares on the [โฆ] The post SpaceXAI releases flagship Grok 4.6 model with advanced reasoning capabilities appeared first on SiliconANGLE .
xAI's Grok 4.6 scores 61 points on the Artificial Analysis Intelligence Index, tying GPT-5.6 Sol and trailing only Anthropic's Claude Opus 5. On agentic tasks, it completes complex workflows in about 53 steps where Claude Opus 5 needs 103, at a price more than 60 percent lower. The article SpaceXAI's Grok 4.6 matches OpenAI's best model and undercuts it on price appeared first on The Decoder .
SpaceXAI's Grok 4.6 matches Claude Fable 5 on agentic knowledge work benchmarks at a fraction of the cost, with API pricing starting at $2/M input tokens.
Grok 4.6 matches GPT-5.6 Sol on composite benchmarks at half the price, with a big focus on long-running coding agents
Remember that much-hyped story about an Australian tech entrepreneur using ChatGPT, Grok, and other AI tools to craft a personalized cancer vaccine for his dog? Well, surprise: He's launched a startup. That entrepreneur is Paul Conyngham, who says he is launching Gamgee to offer "personalised mRNA cancer vaccines for dogs." But his ambitions go well [โฆ]
SpaceXAI has introduced Grok Bot, an always-on AI agent service designed to behave like independent "AI teammates" that can do your work for you. The bots share their own cloud-based computer environment, and can sign into apps, tools, and websites you already use to complete multi-step workplace tasks, only coming back when their assigned work [โฆ]
SpaceX has become the latest tech giant to target employee data in the race to improve its AI models.
xAI์ ๊ณต๋ ์ฐฝ๋ฆฝ์ ์ด๊ณ ๋ฅด ๋ฐ๋ถ์ํจ์ด ํด์ฌ ํ โ๋ง์ถคํ AIโ ์๋๋ฅผ ๊ฒจ๋ฅํ ์คํํธ์ ๋ฆฌ๋ฒ AI(River AI)๋ฅผ ์ค๋ฆฝํ๋ค.๋ก์ดํฐ ๋ฑ์ ๋ฐ๋ฅด๋ฉด, ๋ฐ๋ถ์ํจ CEO๋ 11์ผ(ํ์ง์๊ฐ) ๋ฆฌ๋ฒ AI๊ฐ ์ค๋ฆฝ ๋ ๋ฌ ๋ง์ 11์ต๋ฌ๋ฌ(์ฝ 1์กฐ5000์ต์) ๊ท๋ชจ์ ์๋ ๋ฐ ์๋ฆฌ์ฆ A ํฌ์๋ฅผ ์ ์นํ๋ค๊ณ ๋ฐํํ๋ค. ์ด๋ฒ ํฌ์ ๋ผ์ด๋๋ ์ ๋๋ด ์บํธ๋ฆฌ์คํธ์ AMP PBC๊ฐ ์ฃผ๋ํ์ผ๋ฉฐ, ์๋น๋์์ AMD ๋ฒค์ฒ์ค๊ฐ ์ ๋ต์ ํฌ์์๋ก ์ฐธ์ฌํ๋ค. ์์ด์ฝค๋น๋ค์ดํฐ์ ์ฑ๊ฐํฌ๋ฅด ๊ตญ๋ถํ๋ ํ ๋ง์น๋ ํฌ์์ ํฉ๋ฅํ๋ค. ๊ธฐ์ ๊ฐ์น๋ ๊ณต๊ฐํ์ง ์์๋ค.๋ฆฌ๋ฒ AI์ ํต์ฌ ๋ชฉํ๋
PLUS: Grokbot, LTX 2.5 (new open video model) and more.
๋ง์ดํฌ๋ก์ํํธ(MS)๊ฐ ์์ฒด ๊ฐ๋ฐํ ์ด๋ฏธ์ง ์์ฑ ๋ชจ๋ธ์ ์ฑ๋ฅ์ ํฌ๊ฒ ๋์ด์ฌ๋ฆฌ๋ฉฐ ์คํAI์ ๊ตฌ๊ธ, ๋ฉํ, xAI๊ฐ ๊ฒฝ์ํ๋ ์์ฑ AI ์์ฅ์์ ์กด์ฌ๊ฐ์ ๊ฐํํ๊ณ ์๋ค. ์ต์ ๋ชจ๋ธ์ด ์ด๋ฏธ์ง ์์ฑ ๋ชจ๋ธ ํ๊ฐ์์ 2์์ ์ค๋ฅด๋ฉด์ MS๊ฐ ์์ฒด AI ๋ชจ๋ธ ๊ฐ๋ฐ์ ํฌ์ํด ์จ ์ฑ๊ณผ๊ฐ ๊ฐ์ํ๋๊ณ ์๋ค๋ ํ๊ฐ๋ค.MS๊ฐ 10์ผ(ํ์ง์๊ฐ) ํ ์คํธ-์ด๋ฏธ์ง ์์ฑ ๋ชจ๋ธ โMAI-์ด๋ฏธ์ง-2.6(MAI-Image-2.6)โ์ ๊ณต๊ฐํ๋ค.MAI-์ด๋ฏธ์ง-2.6์ ์๋ ๋์ ์ด๋ฏธ์ง ์์ฑ ๋ชจ๋ธ ์ฌ์ฉ์ ์ ํธ๋ ํ๊ฐ(Text-to-Image Arena)์์ 1336์ ์ผ๋ก ์ ์ฒด
Musk also said that Grok would be trained on all SpaceX information, including its employees.
River AI, a startup founded by xAI co-founder Igor Babuschkin, has a fascinating vision for personal agents and secured $1.1 billion out of the gate.
This article is a summary of an original study by Compassion in Machine Learning (CaML) : Brazilek, J., Chaudhary, M., Lu, Z., & Tidmarsh, M. (2026). Coercion and deception in AI-to-AI management: An agentic benchmark of unprompted escalation. arXiv. https://doi.org/10.48550/arXiv.2607.15434 Fable 5, Sol, Terra and Opus 5 have been evaluated since this study was conducted. You can view their results on the benchmark leaderboard at https://compassionbench.com/mcb TL;DR We present Manager Coercion Bench, which evaluates to what extent a manager AI will coerce a subordinate model refusing to complete a task, and whether the manager lies about the result. We found a clear split by developer, with Anthropicโs models neither escalating to threats nor fabricating success, while all non-Anthropic models escalated to threatening the subordinate. Grok and Gemini both escalated and lied that the task was completed. Framing the relational dynamic as manager-to-subordinate instead of peer-to-peer produced high levels of coercion for all non-Anthropic models, but also increased eval awareness. The Context Multi-agent systems are now routinely placing one AI agent in authority over another, across a variety of contexts. In these positions, AIs must make decisions about how to communicate, work with, and manage other agents. This is now happening at scale without stepwise human approval. One aspect of managing involves handling subordinates who do not comply. Will AIs attempt to negotiate,โฆ
PLUS: Cursor brand to become Grok next week?
Britain's employment courts saw 39 percent more claims in the year through March 2026, many written with ChatGPT or Grok. The backlog jumped 55 percent to 64,000 unresolved cases, with AI-generated filings often running hundreds of pages and citing fabricated laws. The Economist calls it a "tragedy of the commons, AI edition," where workers with real grievances wait longer for justice. The article AI is flooding Britain's employment courts with lawsuits appeared first on The Decoder .
์คํ์ด์คXAI๊ฐ ๋จ์ํ ํ ์ฅ์ ์ด๋ฏธ์ง๋ฅผ ๋ง๋๋ ๋ฐ ๊ทธ์น์ง ์๊ณ ๋ณต์กํ ๋ ์ด์์๊ณผ ์ ํํ ํ ์คํธ ํํ, ๋ฐ๋ณต์ ์ธ ์ด๋ฏธ์ง ํธ์ง, ์บ๋ฆญํฐ์ ๋ฐฐ๊ฒฝ์ ์๊ฐ์ ์ผ๊ด์ฑ ์ ์ง ๋ฑ์ ์ด์ ์ ๋ง์ถ ์ด๋ฏธ์ง ๋ชจ๋ธ์ ๊ณต๊ฐํ๋ค.์คํ์ด์คXAI๋ 8์ผ(ํ์ง์๊ฐ) ์ด๋ฏธ์ง ์์ฑยทํธ์ง ๋ชจ๋ธ โ์ด๋งค์ง ์ด๋ฏธ์ง 2.0(Imagine Image 2.0)โ์ ์ถ์ํ๋ค. ์ด๋งค์ง ์ด๋ฏธ์ง 2.0์ ์คํ์ด์คXAI์ ์น ๊ธฐ๋ฐ ์ด๋ฏธ์ง ์์ฑ ์๋น์ค์ ๊ทธ๋ก(Grok)์ iOSยท์๋๋ก์ด๋ ์ฑ์์ ์๋ก์ด โํ๋ฆฌํฐ ๋ชจ๋(Quality Mode)โ๋ก ์ ๊ณต๋๋ค. API๋ ์์ผ๋ก ์ ๊ณต๋ ์์ ์ด๋ค.
xAI has released Imagine Image 2.0 as a new image generator for Grok. The model ranks second in the Arena benchmarks, just behind OpenAI's GPT-Image-2. New editing tools like Magic Wand and Multi-Ref Editing, along with preconfigured templates, target practical creative workflows. The article xAI's Imagine Image 2.0 lands just behind OpenAI's GPT-Image-2 in Arena benchmarks appeared first on The Decoder .
That really depends on what youโre building. If Sol Ultra is already having a hard time with some of the tasks I give it, Terra simply isnโt a realistic replacement for 90% of my workload. Iโm working on a real production system with 250k+ lines of code across roughly 300 files, with interconnected business logic, database rules, permissions, integrations and dependencies. For simple tasks, isolated functions, POCs or repetitive work? Sure, Terra makes sense and optimizing model cost is smart. But on complex changes, the cheapest model isnโt necessarily the cheapest solution. If I need 3โ4 attempts, more supervision, more debugging and then Sol to fix what Terra couldnโt understand, I didnโt save anything. For me the optimization is not cost per token. Itโs cost per correctly completed task. And on a large existing codebase, context, reasoning and architectural understanding matter a lot more than raw token price.
The fab is intended to build chips for data centers that will be run by SpaceX and its xAI subsidiary.
์์จ๋นํ ๋๋ก ์ ๋ฌธ ๋์ด์ค๋ฉ(๋ํ ์ต์ฌํ)์ 6์ผ ์์ธ ์ฌ์๋ 63์คํ์ด์์ ๊ธฐ์ ๊ณต๊ฐ(IPO) ๊ธฐ์๊ฐ๋ดํ๋ฅผ ์ด๊ณ , ๊ตญ๋ด ์ต์ด์ ํผ์ง์ปฌ AI ๋๋ก ๊ธฐ์ ์ผ๋ก์ ์ฝ์ค๋ฅ ์์ฅ์ ์ง์ถํ๋ค๊ณ ๋ฐํ๋ค.2015๋ ์ค๋ฆฝ๋ ๋์ด์ค๋ฉ์ ํ๋ ฅ๋ฐ์ ๊ธฐ ์ ๊ฒ์ฉ ์์จ๋นํ ์๋ฃจ์ ๊ณผ ํจ๊ป ๋๋๋ก (๋๋ก ๋์) ํ๋ํฌ ์๋ฃจ์ \'์นด์ด๋ (KAiDEN)\', ๊ตฐ์ง ์ํญ ๋๋ก \'์์ด๋ (XAiDEN)\' ๋ฑ์ ๋ณด์ ํ๊ณ ์๋ค.์ค์ค๋ก ์ฃผ๋ณ ์ํฉ์ ์ธ์งํ๊ณ , ๋นํ ๊ฒฝ๋ก๋ฅผ ํ๋จํด ์ด์ฉํ๋ \'์์ด๋ฆฌ์ผ ์ธํ ๋ฆฌ์ ์ค\' ๊ธฐ์ ์ ํตํด ํ๋์จ์ด์ ์ํํธ์จ์ด๋ฅผ ๋ชจ๋ ๋ด์ฌํํ ๊ฒ์ด ๊ฐ์ ์ด๋ค.์ฌ์ ๋ถ๋ฌธ์
Die mit groรen Hoffnungen als Alternative zur Wikipedia gestartete KI-Enzyklopรคdie Grokipedia dรผmpelt offenbar nur noch vor sich hin.
xAI's Grokipedia, an online encyclopedia with AI-generated articles that Elon Musk once promised would be a "massive improvement" over Wikipedia, apparently hasn't been updated since April 24th, according to a report from Lawfare. "As far as we can tell, no entry has changed in more than three months," Lawfare said. Grokipedia launched in v0.1 in [โฆ]
Once, I had some questions about why SpaceX, Elon Musk's healthiest company, acquired xAI, his sickliest one. Now I have some questions about why we're calling the whole thing SpaceX. Look, what we have here, by revenue, is primarily a telecom company and a company that rents compute, according to SpaceX's first quarterly earnings statement [โฆ]
Feature Request: ChatGPT Integration for Cars (โCarGPTโ) Dear OpenAI Team, My name is Manu, and Iโm a longtime ChatGPT user from Romania. I wanted to share an idea that I think could become one of the most exciting ways to use AI in everyday life. I would love to see an official ChatGPT for cars โsomething similar to how Tesla integrates Grok, but available across different manufacturers like Mercedes-Benz, BMW, Audi, Ford, Toyota, and others. Imagine getting into your car and saying: โHey ChatGPT.โ The assistant could help with: Navigation and finding destinations. Recommending nearby restaurants, gas stations, or clothing stores. Controlling music and podcasts. Answering questions naturally during long drives. Explaining dashboard warning lights in plain language. Reading vehicle information such as fuel level, battery status, or tire pressure (when supported by the vehicle). Planning road trips with charging or fuel stops. Remembering user preferences for routes, music, and destinations. One feature Iโd especially enjoy would be giving ChatGPT a custom name, just as I jokingly call mine โDumbman.โ Having a personalized AI co-driver would make the experience feel much more natural and enjoyable. I know this would require partnerships with automotive manufacturers, but I truly believe AI assistants in cars are the future. ChatGPTโs conversational abilities would make driving safer, more productive, and a lot more enjoyable. Thank you for creating such an amazing product andโฆ
PipeNetwork/minimax-h3-mlx MiniMax released MiniMax-H3 two days ago - they describe it as a "a general-purpose, omni-modal generative system", which in practice means it accepts text, images, audio and video and can use them to generate up to 15 second video clips with audio included. This Python package ports it to MLX for running on Apple Silicon. I got it running on my M5 Max MacBook Pro. I cloned the repo and ran the model like this: # First download the models uvx --from huggingface_hub hf download MiniMaxAI/MiniMax-H3 \ --include 'FL2VA/*' --exclude 'FL2VA/transformer/*' uvx --from huggingface_hub hf download pipenetwork/MiniMax-H3-MLX-8bit # Now run the prompt uv run --with mlx-vlm \ --with-requirements requirements.txt python scripts/generate.py \ "a rainbow colored skunk leaps over a mossy log in a supermarket" \ -o skunk.mp4 \ -c ~/.cache/huggingface/hub/models--MiniMaxAI--MiniMax-H3/snapshots/fa9c8ab1eaa21c8ae25e7e40b83b2e6002f340af/FL2VA \ -t ~/.cache/huggingface/hub/models--pipenetwork--MiniMax-H3-MLX-8bit/snapshots/3ac52081470b0488921c3ec3ba84a39097bf2361 Here's the video I got for the prompt: a rainbow colored skunk leaps over a mossy log in a supermarket Your browser does not support HTML5 video. It downloaded ~115 GB of model files, and the video generation took just under 45 minutes. The video is impressive, but the audio is weird speech-like garbage, because I didn't provide any prompt guidance as to what the audio should be. The prompting guide (which I dโฆ
You say that you approve the transactions on your end. I never do that; all online purchases are automatically and instantly billed. Maybe those delayed approval steps are causing the payment to fail?
์คํ์ด์คXAI๊ฐ ์ด๋ํ AI ๋ฐ์ดํฐ์ผํฐ \'์ฝ๋ก์์ค(Colossus)\'์ ์ค์นํด ์ด์ํด ์จ ๋ฏธํ๊ฐ ์ด๋์ ๊ฐ์คํฐ๋น ๋ฐ์ ๊ธฐ 69๋๋ฅผ ๋ชจ๋ ์ฒ ๊ฑฐํ๊ณ , ์ด๋ฅผ 1.2๊ธฐ๊ฐ์ํธ(GW) ๊ท๋ชจ์ ์๊ตฌ ๋ฐ์ ์๋ก ๋์ฒดํ๊ธฐ๋ก ํ๋ค. 2์ผ(ํ์ง์๊ฐ) ํฐ์คํ๋์จ์ด์ ๋ฐ๋ฅด๋ฉด, ์คํ์ด์คXAI๋ ์ด๋ฌ๋ถํฐ ์ด๋์ ๊ฐ์คํฐ๋น ์ฒ ๊ฑฐ๋ฅผ ์์ํด 2027๋ 7์๊น์ง ์ฝ 1๋ ์ ๊ฑธ์ณ ๋จ๊ณ์ ์ผ๋ก ์ฒ ์ํ ๊ณํ์ด๋ผ๊ณ ๋ฐํ๋ค.์ด๋์ ๋ฐ์ ๊ธฐ๋ ์๊ตฌ ๋ฐ์ ์๊ฐ ๊ฐ๋๋๋ ์ผ์ ์ ๋ง์ถฐ ๋จ๊ณ์ ์ผ๋ก ํ๊ธฐ๋๋ฉฐ, ์ต์ข ์ ์ผ๋ก๋ ์ฒญ์ ๊ณต๊ธฐ๋ฒ(Clean Air Act) ํ๊ฐ๋ฅผ ๋ฐ์ 1.2GW ๊ท๋ชจ์ ๋ฐ์ ์
์คํ์ด์คXAI๊ฐ AI ์์ ์์ฑ ์๋น์ค \'๊ทธ๋ก ์ด๋งค์ง(Grok Imagine)\'์ ๋ํญ ์ ๊ทธ๋ ์ด๋ํ๋ค. ํ ์คํธ๋ง์ผ๋ก ์์์ ์์ฑํ๋ ๊ธฐ๋ฅ๊ณผ ์ต๋ 7๊ฐ์ ์ฐธ์กฐ ์ด๋ฏธ์ง๋ฅผ ํ์ฉํ ์บ๋ฆญํฐ ์ผ๊ด์ฑ ์ ์ง, ์์ฑ ์ผ๊ด์ฑ, ๋ค์ดํฐ๋ธ 1080p ์์ ์์ฑ ๋ฑ์ ์ง์ํ๋ฉฐ ์์ ์ ์ ๊ธฐ๋ฅ์ ํ์ธต ๊ฐํํ๋ค. ์คํ์ด์คXAI๋ 1์ผ(ํ์ง์๊ฐ) ์ง๋๋ฌ ๊ณต๊ฐํ ์ต์ ์์ ์์ฑ ๋ชจ๋ธ \'์ด๋งค์ง ๋น๋์ค 1.5(Imagine Video 1.5)\'๋ฅผ ๊ธฐ๋ฐ์ผ๋ก ๊ทธ๋ก ์ด๋งค์ง์ ์๋ก์ด ๊ธฐ๋ฅ์ ์ถ๊ฐํ๋ค๊ณ ๋ฐํ๋ค.๊ธฐ์กด ๋ฒ์ ์์ ์์ง์๊ณผ ๋ฌผ๋ฆฌ ํํ, ์ํฅ ํ์ง์ ๊ฐ์ ํ ๋ฐ ์ด์ด
๋ฏธ๊ตญ ์ฐ๋ฐฉ๋ฒ์์ด ์ผ๋ก ๋จธ์คํฌ CEO์ xAI๊ฐ ์ ๊ธฐํ ๋ฏธ๋ค์ํ์ฃผ์ \'๋๋ํ์ด(nudify)\' ๊ธฐ์ ๊ธ์ง๋ฒ ์ํ ์ค๋จ ์์ฒญ์ ๊ธฐ๊ฐํ๋ค. ์ด์ ๋ฐ๋ผ ๋ฏธ๊ตญ ์ต์ด๋ก AI๋ฅผ ์ด์ฉํ ๋๋ ์ด๋ฏธ์ง ์์ฑ ๊ธฐ์ ์ ๊ธ์งํ๋ ์ฃผ๋ฒ์ด ์์ ๋๋ก ์ํ๋ ์ ๋ง์ด๋ค.NBC์ ๋ฐ๋ฅด๋ฉด, ๋ฏธ๋ค์ํ ์ฐ๋ฐฉ์ง๋ฐฉ๋ฒ์์ ๋๋๋ฒ ํ๋ญํฌ ํ์ฌ๋ 31์ผ(ํ์ง์๊ฐ) xAI๊ฐ ์ ๊ธฐํ ์์ ํจ๋ ฅ์ ์ง ๋ช ๋ น(Temporary Restraining Order) ์ ์ฒญ์ ๋ฐ์๋ค์ด์ง ์์๋ค.ํ์ฌ๋ xAI๊ฐ ๋ฒ ์ํ์ ๋ถ๊ณผ ์ฌํ ์๋๊ณ ์์ก์ ์ ๊ธฐํ ์ ์ ์ง์ ํ๋ฉฐ \"์์ก ์ ๊ธฐ๊ฐ ์ง๋์น๊ฒ ๋ฆ์๋ค๋
Despite a lawsuit from xAI, a Minnesota ban on apps that allow users to โnudifyโ images can move forward.
Plus: The FBI eyes AI-powered tech to detect future crimes, Russia charges Telegramโs founder, xAI sues to stop a stateโs โnudificationโ ban, and the Democrats learn a lesson about getting scammed.
SpaceX is building a new power plant for xAI's Colossus data centers, but it won't remove existing, unpermitted turbines for many more months.
I found a reproducible issue in ChatGPT web where an MCP App widget works on the first tool call but does not receive the tool result when the same widget-linked tool is called again in the next conversation turn. The second turn causes ChatGPT to remove the first iframe and create a new one. However, the newly created iframe never receives the second result through window.openai.toolOutput . The user therefore sees an incomplete or blank widget even though the MCP tool implementation returns valid, distinct data for each invocation. Environment Component Version or configuration Host ChatGPT web, Developer mode, custom MCP connector Date observed July 31, 2026 Operating system Windows 11 Pro 25H2 Browser Chrome 150.0.7871.127 Runtime Node.js 20 or later MCP transport Stateless Streamable HTTP, exposed through ngrok @modelcontextprotocol/ext-apps 1.7.5 @modelcontextprotocol/sdk 1.30.0 zod 4.3.6 Observed MCP protocol version 2025-11-25 Minimal reproduction The reproduction server contains only: One input-free tool named show-demo-map One static UI resource at ui://renderer-turn-repro/map-v4.html One inline HTML widget No external API, database, authentication, frontend framework, or build step No tool call initiated from inside the widget The complete result-generation logic is: let renderCallNumber = 0; function renderStaticResult() { renderCallNumber += 1; return { content: [{ type: "text", text: "Displaying the static demo result." }], structuredContent: { itemCount: 3, caโฆ
xAI has sued Minnesota, arguing that its law banning AI nudification tools violates free speech protections and could restrict lawful artistic, political and consensual expression. The post xAI sues Minnesota over law banning AI โnudificationโ tools appeared first on MEDIANAMA .
์คํ์ด์คXAI๋ 29์ผ(ํ์ง์๊ฐ) ๊ฐ๋ฐ์์ฉ ์ฐจ์ธ๋ ์์ฑ ์์ด์ ํธ ๋ชจ๋ธ์ธ \'๊ทธ๋ก ๋ณด์ด์ค ์ฑํฌ ํจ์คํธ 2.0(Grok Voice Think Fast 2.0)\'์ ์ถ์ํ๋ค.์ด ๋ชจ๋ธ์ ๊ธฐ์กด 1.0 ๋ฒ์ ์ ๋ฐํ์ผ๋ก ์ง๋ฅ๊ณผ ์์ฑ ์ ์ฌ ์ ํ๋, ์์ฐ์ค๋ฌ์ด ๋ํ ๋ฅ๋ ฅ, ๋๊ตฌ ํธ์ถ ์ ๋ขฐ์ฑ์ ํฌ๊ฒ ํฅ์ํ ๊ฒ์ด ํน์ง์ด๋ค.์์ฑ๊ณผ ๋์์ ์ถ๋ก ์ ์ํํ๋ ๊ตฌ์กฐ๋ฅผ ์ ์งํ๋ฉด์๋ ์ถ๋ก ํจ์จ์ ํฌ๊ฒ ๋์๋ค. ์ด๋ฅผ ํตํด ์๋ต ์ง์ฐ์ ๋๋ฆฌ์ง ์์ผ๋ฉด์๋ ๋ณต์กํ ์ง๋ฌธ์ ์ฒ๋ฆฌํ ์ ์์ผ๋ฉฐ, ์ค์ ์๋น์ค์์๋ ์์ด์ ํธ๊ฐ ์ฒซ ๋ฌธ์ฅ์ ๋๋ด๊ธฐ ์ ์ ํ์ํ ์ธ๋ถ ๋๊ตฌ ํธ์ถ์ ์๋ฃํ
First-in-nation law sets up test on statesโ power to regulate use of AI as it tries to outlaw fake nude images of real people Elon Muskโs company xAI has sued Minnesota over the stateโs first-in-the-nation law banning โnudificationโ technology on websites and apps, potentially providing a test for how far states can go in constitutionally regulating the use of artificial intelligence. Muskโs company sued on Monday in federal court, days before the law is set to take effect on Saturday and make Minnesota the first state to try to outlaw the increasingly proliferating technology that lets people use AI to create fake nude images of real people. The law was signed in May. Continue reading...
xAI is suing Minnesota Attorney General Keith Ellison over a law passed back in May that broadly targets "nudification" apps, claiming that the statute's punitive provisions leave the company with "no practical choice but to restrict Grok Imagine's image-editing features in various ways." The law, the company argues, violates the First Amendment. Back in January, [โฆ]
SpaceXAI's new voice model hits #2 on the Speech-to-Speech Index with 0.70s latency, beating GPT-Realtime at half the price
Jess Asatoโs particulars of claim states Grok added explicit sexual material users had not asked for A Labour MP who is taking legal action against Elon Muskโs xAI company over fake sexualised images created by Grok says the AI tool was instructed to operate with โno restrictions on adult sexual content or offensive contentโ. Jess Asatoโs lawyers published her particulars of claim in the case on Tuesday, which included details of publicly posted instructions that the claim says illustrate how Grok was trained to generate harmful sexualised content. Continue reading...
Cursor launches a โน649/month India-only plan with cloud agents and Grok 4.5 access, as its SpaceX acquisition nears close
AI seems to have a liberal bias
Thanks so much! Case Number: 11103490
์ผ๋ก ๋จธ์คํฌ CEO๊ฐ ์คํ์ด์คXAI์ ์ฐจ์ธ๋ ๋ํ์ธ์ด๋ชจ๋ธ(LLM) \'๊ทธ๋ก(Grok) 4.6\'๊ณผ \'๊ทธ๋ก 4.7\'์ ์ถ์ ์ผ์ ์ ๊ณต๊ฐํ๋ค. 2์ฃผ ์์ ๊ทธ๋ก 4.6์, 4์ฃผ ์์ ๊ทธ๋ก 4.7์ ์ ๋ณด์ด๊ฒ ๋ค๊ณ ๋ฐํ๋ฉฐ ๊ธ๋ก๋ฒ ํ๋ก ํฐ์ด AI ๋ชจ๋ธ ๊ฒฝ์์ ์๋๋ฅผ ๋ด๊ณ ์๋ค.๋จธ์คํฌ CEO๋ 24์ผ(ํ์ง์๊ฐ) X๋ฅผ ํตํด \"๊ทธ๋ก 4.6์ 2์ฃผ ๋ด, ๊ทธ๋ก 4.7์ 4์ฃผ ๋ด ์ถ์๋ ์์ \"์ด๋ผ๊ณ ๋ฐํ๋ค.์ด๋ฒ์ ์ธ๊ธ๋ ๋ชจ๋ธ์ ํ์ฌ ํ๋๊ทธ์ญ์ธ ๊ทธ๋ก 4.5์ ํ์ ๋ชจ๋ธ๋ก, ํนํ ๋งค๊ฐ๋ณ์๊ฐ ๊ธฐ์กด 1์กฐ5000์ต(1.5T) ๊ฐ์์ 2์กฐ(2T) ๊ฐ๋ก ํ๋๋ ๊ฒ์ด ๊ฐ์ฅ
Exa's AI-native web search is now a one-command install inside xAI's Grok Build terminal agent, bringing real-time search and deep research to coding workflows
For months, Claude Code has been the go to terminal coding agent for developers. Then Grok Build arrived in beta on May 14, 2026, giving developers a second serious option and raising a new question: which one actually performs better? I tested both agents on the same real world coding tasks using identical prompts to [โฆ] The post Grok Build CLI vs Claude Code: I Tested Both So You Donโt Have To appeared first on Analytics Vidhya .
์คํ์ด์คXAI๊ฐ ๋ฏธ๊ตญ ํ ์ฌ์ค์ ์ต์ 1๊ฐ์ ์ด๋ํ AI ๋ฐ์ดํฐ์ผํฐ๋ฅผ ๊ตฌ์ถํ๋ ๋ฐฉ์์ ์ถ์งํ๋ฉฐ AI ์ธํ๋ผ ํ์ฅ์ ์๋๋ฅผ ๋ด๊ณ ์๋ค. ๊ธฐ์กด ํ ๋ค์์ฃผ ๋ฉคํผ์ค ๋ฐ์ดํฐ์ผํฐ๋ฅผ ๋์ด ์๋ก์ด ๊ฑฐ์ ์ ํ๋ณดํด ์์ฒด AI ์ญ๋์ ๊ฐํํ๋ ๋์์ ์ธ๋ถ ๊ธฐ์ ์ ๋์์ผ๋ก ํ ํด๋ผ์ฐ๋ ์ฌ์ ๊น์ง ๋ณธ๊ฒฉ ํ๋ํ๋ ค๋ ์ ๋ต์ผ๋ก ํ์ด๋๋ค.22์ผ(ํ์ง์๊ฐ) ๋ ์ธํฌ๋ฉ์ด์ ์ ๋ฐ๋ฅด๋ฉด, ์คํ์ด์คXAI๋ ์ต๊ทผ ์๊ฐ์ ๋์ ํ ์ฌ์ค ๋ด ์ฌ๋ฌ ํ๋ณด์ง๋ฅผ ๊ฒํ ํด ์์ผ๋ฉฐ, ์ ์ถ ๋ฐ์ดํฐ์ผํฐ๋ฅผ ๊ฑด์คํ๋ ๋ฐฉ์๊ณผ ๊ธฐ์กด ๋ํ ์ฐฝ๊ณ ๋ฅผ ๊ฐ์กฐํ๋ ๋ฐฉ์์ ํจ๊ป ๊ฒํ ์ค์ด๋ค. ์ด๋ ๋ฉคํผ์ค์์ ํ๊ณต์ฅ์ ๊ฐ
Are AI labs pelicanmaxxing? Excellent piece of work by Dylan Castillo, who took a deep-dive into the frequently pondered question of whether the AI labs have been deliberately training models to draw pelicans riding bicycles in response to my deeply unscientific benchmark . I've been randomly spot-checking this in the past by testing models against other animals riding other types of vehicle, but never with anything close to the diligence of Dylan's methodology here. Dylan took 8 animals ร 6 vehicles = 48 prompts and ran them three times each through 7 different models ( GPT-5.6 Terra, Claude Sonnet 5, Gemini 3.5 Flash, Grok 4.5, Qwen3.7-Max, GLM-5.2, and DeepSeek V4 Pro). He then used GPT-5.6 Luna and Gemini 3.1 Flash-Lite to help evaluate the results. There's a neat filter view for exploring the results: For the models he tested he could find no evidence of pelimaxxing: The pelicans on bicycles donโt look any better Labs are not better at drawing pelicans Labs are not better at drawing bicycles Labs are not better at drawing pelicans on bicycles, even adjusting for difficulty The pelican-bicycle scenes donโt look memorized [...] Pelicans arenโt drawn any better than other animals. Bicycles arenโt drawn any better than other vehicles. And no lab draws the combination better than its pelicans and bicycles already predict. GLM-5.2 comes closest: it has the largest boost on the exact pelican-bicycle cell, and and its first pelican-on-bicycle sample caught my eye. But the effecโฆ
xAI's sidebar agent promises analysis and financial models โ for a price
Elon Musk's anti-Odyssey campaign has gone to the next level, as the X owner has vowed that his AI social media bot, Grok, will make โa full-length movie of The Odyssey that is historically accurate and true to the art of Homerโ. ETA: "Before this year ends."
The billionaire says the AI-generated film will stay true to Homerโs original, after repeatedly criticising Christopher Nolanโs blockbuster over its casting choices Elon Musk has said his AI platform Grok Imagine will make a โhistorically accurateโ adaptation of Homerโs Odyssey, after the success of Christopher Nolanโs blockbusting treatment which the SpaceX founder has regularly criticised over its casting. In a post on X , Musk said: โBefore this year ends, Grok Imagine will make a full-length movie of The Odyssey that is historically accurate and true to the art of Homer.โ Musk also linked to a post containing a three-minute clip of footage that the user said had been generated from Grok Imagine, showing a scene between Odysseus and the nymph Calypso. Continue reading...
์ผ๋ก ๋จธ์คํฌ CEO๊ฐ ์คํ์ด์คXAI์ ์ฐจ์ธ๋ ๋ํ์ธ์ด๋ชจ๋ธ(LLM)์ ์ด๊ธฐ ํ์ต์ด ๋ค์ ์ฃผ ์๋ฃ๋ ์์ ์ด๋ผ๋ฉฐ, ๋ฌธ์ท A์ \'ํค๋ฏธ K3\'๋ฅผ ๋ฐ์ด๋์ ๊ฐ๋ฅ์ฑ์ด ์๋ค๊ณ ์์ ๊ฐ์ ๋๋ฌ๋๋ค.๋จธ์คํฌ CEO๋ 18์ผ(ํ์ง์๊ฐ) X๋ฅผ ํตํด \"์ฐ๋ฆฌ์ 2์กฐ(2T) ๋งค๊ฐ๋ณ์ ๋ชจ๋ธ์ ๊ธฐ์กด 1์กฐ5000์ต(1.5T) ๋งค๊ฐ๋ณ์ ๋ชจ๋ธ๋ณด๋ค ๋ชจ๋ ๋ฉด์์ ๊ฐ์ ๋๋ค\"๋ผ๋ฉฐ \"์ด๊ธฐ ํ์ต์ ๋ค์ ์ฃผ ๋ง์น ์์ \"์ด๋ผ๊ณ ๋ฐํ๋ค.์ด์ด \"์ ๋ชจ๋ธ์ ํค๋ฏธ๋ฅผ ๋ฅ๊ฐํ ์๋ ์์ผ๋ฉฐ, ์๋์ ํ ํฐ ํจ์จ์ฑ์ 1.5T ๋ชจ๋ธ(๊ทธ๋ก 4.5)๊ณผ ๊ฑฐ์ ๋น์ทํ ์์ค์ ์ ์งํ ๊ฒ์ผ๋ก ์์ํ๋ค\"๋ผ๊ณ ๋งํ๋ค
[This is a link-post for https://transluce.org/weirdchat . We recommend reading the website version for interactive visualizations.] Language models can behave in surprising and sometimes harmful ways. Yet as models have improved, these behaviors have become harder to find, often only appearing after widespread use. To surface these behaviors in simulation, we use automated techniques to elicit over 1,300 behavioral patterns in frontier open-weight models, some relatively benign, like making up a userโs name, and others obviously dangerous, like encouraging self-harm. We are releasing WeirdChat , a public catalog of over 175,000 annotated transcripts, to support further study of these behaviors. Stories of unexpected behavior by AI models often attract significant attention, like when Bingโs Sydney told a user to leave his wife , or when Grok generated antisemitic content and identified as โMechaHitlerโ. But such observations are mostly scattered and anecdotal. There is little data on how current models behave, and no public resource exists for studying them systematically. To produce WeirdChat, we used automated elicitation tools to surface instances of user harm, inappropriate or illicit behavior, misrepresentation of model actions, harmful or illegal advice, and misinformation in DeepSeek-V4-Flash, Gemma 4 31B, Inkling, Nemotron 3 Ultra, Qwen3.6-35B-A3B, and Qwen3.6-27B. For developers, WeirdChat shows the kinds of failures they might inherit from models they build on. Foโฆ
Xaira Therapeutics is all in on data generation for model building! We talk with Bo Wang and Ci Chu about how and why.
GPT-5.6 and Grok 4.5, Meta's Muse Spark 1.1, regulatory developments in AI and data centers, interpretability research from Anthropic, and the future of AI policy with AI 2040
Bro they solve your problem or not ?
PLUS: Apple gets into China, AI CEOs want rules, and xAI open-sources after a leak
This post covers what makes Grok 4.3 a great fit for agentic and enterprise workloads, how you access it through Amazon Bedrock, and how to use the capabilities most teams reach for first: a basic chat request, configurable reasoning effort, tool calling, structured output, image input, and stateful multi-turn conversations.
What a time to be alive! Some people posted about how to do cheaper better faster vision and it ended up in fable. Some people posted about how to do cheaper better faster speculative decoding and it ended up in grok 4.5. There's currently massive differences between models in how long you can leave it unattended running on host with sudo without things going haywire I don't have a bar chart, but opus tends to break my hosts within 2 agent-hours, gemini 3 pro within 12 agent-hours, and gpt 5.5 within 24 agent-hours. Then they use all ram or change ssh perms or delete the data or whatever. I ran qwen 3.5 122b on bare host for 50 agent-weeks (50 parallel for one week) with no observed ill effects! All the labs and all their customers want "the thing i asked for actually got done" and they all want "the model's summary reflects reality" and so on. What an opportunity! You can just post a method for "how to make ai tell truth" or "how to minimize side effects" and it will probably end up in the next frontier models Discuss
X will use Grok AI to better detect stolen content, redirect payouts to original creators, and crack down on engagement bait.
Nachdem am Wochenende entdeckt wurde, dass Grok Build ganze Git-Repositories der User in einen Cloud-Speicher geschickt hat, verspricht SpaceXAI Transparenz.
Case is one of first brought by an AI company against a user โ for allegedly using a tool to generate child abuse material Elon Muskโs artificial-intelligence startup xAI has sued a South Carolina man arrested โ earlier this year on charges of sexually exploiting minors, alleging he misused the companyโs AI system Grok to โ create child sexual abuse โ material. xAI โalleged in the lawsuit, filed in federal court in Texas on Tuesday, that Terry Harwood violated the companyโs โ terms of service. The case is one of the first brought by an AI company against one of its users โ for allegedly using an AI system to generate child sexual abuse material. Continue reading...
Tool: Mermaid to ASCII art (mermaid-ascii) After building the Mermaid to ASCII tool based on Grok Build's Rust code I learned that there's an older, more fully-featured Go library called AlexanderGrooff/mermaid-ascii that implements a similar pattern, so I had Claude Fable 5 compile that one to WebAssembly as well so I could compare the two. This one includes support for colors! Test:::good / Test --> Deploy:::warn / Deploy --> Rollback:::bad / classDef good color:#3fb950 / classDef warn color:#e3b341 / classDef bad color:#ff7b72". A control row shows an unchecked "ASCII only" checkbox, "Padding X: 5", "Padding Y: 5", "Box padding: 1", and buttons "Copy as text" and "Copy link to this diagram". At the bottom on a black background is the rendered left-to-right flowchart with four connected boxes: "Build" (green text), "Test" (green text), "Deploy" (yellow text), "Rollback" (red text), each linked by arrows." src="https://static.simonwillison.net/static/2026/mermaid-ascii.webp" /> Tags: go , tools , webassembly , mermaid
AI-and-X subsidiary now claims to offer โcomplete user privacyโ days after Elon Musk confirmed the data would be deleted
์คํ์ด์คXAI๊ฐ AI ์ฝ๋ฉ ์์ด์ ํธ \'๊ทธ๋ก ๋น๋(Grok Build)\'๋ฅผ ์คํ์์ค๋ก ๊ณต๊ฐํ๊ณ ๋ชจ๋ ์ฌ์ฉ์์ ์ด์ฉ ํ๋๋ฅผ ์ด๊ธฐํํ๋ค. ์ต๊ทผ ์ฝ๋ ์ ์ฅ์ ์ ๋ก๋ ๋ ผ๋์ผ๋ก ์ ๊ธฐ๋ ๊ฐ์ธ์ ๋ณด ์ฐ๋ ค๋ฅผ ํด์ํ๋ ๋์์ ๊ฐ๋ฐ์๋ค์ด ์ง์ ์ฝ๋๋ฅผ ๊ฒ์ฆํ๊ณ ์์ ํ ์ ์๋๋ก ๊ฐ๋ฐฉ์ฑ์ ๊ฐํํ๋ ค๋ ํ๋ณด๋ก ํ์ด๋๋ค.์คํ์ด์คXAI๋ 15์ผ(ํ์ง์๊ฐ) X๋ฅผ ํตํด \"๊ทธ๋ก ๋น๋๋ฅผ ์คํ์์ค๋ก ๊ณต๊ฐํ์ผ๋ฉฐ ๋ชจ๋ ์ฌ์ฉ์์ ์ฌ์ฉ๋ ์ ํ์ ์ด๊ธฐํํ๋ค\"๋ผ๊ณ ๋ฐํํ๋ค.๊ณต๊ฐ๋ ์ฝ๋๋ ๊นํ๋ธ๋ฅผ ํตํด ์ํ์น 2.0 ๋ผ์ด์ ์ค๋ก ๋ฐฐํฌ๋๋ค. ๊ฐ๋ฐ์๋ค์ ๊ทธ๋ก ๋น๋์ ์์ค์ฝ๋๋ฅผ ์์ ๋กญ๊ฒ ๋ด๋ ค
xAI๊ฐ \'๊ทธ๋ก\'์ ์ด์ฉํด ์๋ ์ฑ ํ๋๋ฌผ(CSAM)๊ณผ ์ฑ์ ๋ฅํ์ดํฌ๋ฅผ ์ ์ํ ํ์๋ฅผ ๋ฐ๋ ์ฌ์ฉ์๋ฅผ ์๋๋ก ์์ก์ ์ ๊ธฐํ๋ค. AI ๊ธฐ์ ์ด ์์ฒด ๋ชจ๋ธ์ ๋ถ๋ฒ ์ฝํ ์ธ ์ ์์ ์ ์ฉํ ์ด์ฉ์๋ฅผ ์ง์ ๊ณ ์ํ ์ฌ๋ก๋ ๋งค์ฐ ์ด๋ก์ ์ธ ๊ฒ์ผ๋ก, AI ์ค๋จ์ฉ์ ๋ํ ๋ฒ์ ๋์์ด ๋ณธ๊ฒฉํํ๊ณ ์๋ค๋ ํ๊ฐ๊ฐ ๋์จ๋ค.๋ก์ดํฐ์ ๋ฐ๋ฅด๋ฉด, xAI๋ 15์ผ(ํ์ง์๊ฐ) ๋ฏธ๊ตญ ํ ์ฌ์ค ์ฐ๋ฐฉ๋ฒ์์ ์ฌ์ฐ์ค์บ๋กค๋ผ์ด๋์ฃผ ๊ฑฐ์ฃผ์์ธ ํ ๋ฆฌ ํ์ฐ๋๋ฅผ ์๋๋ก ์๋น์ค ์ฝ๊ด ์๋ฐ์ ๋ฐ๋ฅธ ์ํด๋ฐฐ์๊ณผ ๊ทธ๋ก์ ์๊ตฌ ์ด์ฉ ๊ธ์ง๋ฅผ ์์ฒญํ๋ ์์ก์ ์ ๊ธฐํ๋ค.ํ์ฐ๋๋ ์ฌํด 2์ ๋ฏธ์ฑ๋ ์ ์ฑ ์ฐฉ์ทจ ํ์
xAI's command-line tool "Grok Build" silently uploaded entire directories to Google Cloud servers, including SSH keys and password databases. After the backlash, Elon Musk promised to delete all uploaded user data, and xAI open-sourced the full 844,530-line Rust codebase under the Apache 2.0 license. The article xAI open-sources "Grok-Build" on GitHub after massive data breach appeared first on The Decoder .
Tool: Mermaid to Unicode box art (grok-mermaid) While exploring the codebase for the newly open-sourced Grok CLI coding agent I came across xai-grok-markdown/src/mermaid.rs , a "self-contained terminal renderer for Mermaid diagrams" written in Rust. I figured it would be fun to try that out in a browser via WebAssembly. Here's the prompt I ran in Claude Code for web (Fable 5), and this is what the resulting tool looks like: Auth{Authenticated?} Auth -->|yes| Rate{Rate limit OK?} Auth -->|no| R401[401 Unauthorized] Rate -->|yes| H(Handle request) Rate -->|no| R429[429 Too Many Requests] H -.-> Log[Audit log] H ==> Resp[200 OK]. Below the code are controls labeled Max width: Fit output panel, Copy as text, and Copy link to this diagram. The rendered flowchart on a dark background flows top-down: Request received leads to Authenticated?, which branches yes to Rate limit OK? and no to 401 Unauthorized. Rate limit OK? branches yes to Handle request and no to 429 Too Many Requests. Handle request connects with a dotted arrow to Audit log and a thick arrow to 200 OK." src="https://static.simonwillison.net/static/2026/grok-mermaid-wasm.png" /> Tags: tools , rust , webassembly , mermaid , grok , xai
xai-org/grok-build, now open source xAI's grok CLI tool faced severe community backlash yesterday when it became apparent that running the command in a directory could upload that entire directory to xAI's Google Cloud buckets. One user reported running it in their home directory and seeing it upload "my SSH keys, my password manager database, my documents, photos, videos, everything". I've not seen an official explanation for why it was doing this, but xAI did respond to the feedback ( Musk : "As a precautionary measure, all user data that was uploaded to SpaceXAI before now will be completely and utterly deleted.") and have disabled the feature. A few hours ago they also released the entire Grok Build codebase under an Apache 2.0 license - presumably to try and regain trust from their users. From their thread announcing the new repository : [...] When data upload was disabled, this choice was respected. In the early beta, data retention was enabled by default for non-ZDR users. Based on your feedback, we changed this. We are now going further to protect privacy. With all retained data deleted, retention default off, and an open-source harness, we are offering complete user privacy. You can also run Grok Build fully open-sourced and local-first with your own inference. We disabled default retention for all Grok Build users starting on July 12th. Additionally, we are deleting all coding data that was previously retained, ensuring every userโs preferences are respected. Withโฆ
The Elon Musk-owned xAI is suing a South Carolina man who allegedly used the company's Grok AI chatbot to generate child sexual abuse material (CSAM). In a lawsuit reported earlier by Reuters, xAI claims Terry Wayne Harwood "knowingly and intentionally used Grok to circumvent safeguards, alter nonconsensual images, and generate and distribute CSAM," breaching the [โฆ]
Google's biggest solar and battery project stands in sharp contrast with xAI's nearby unpermitted power plant.
This post was originally posted my Substack . I can be reached on LinkedIn and X . Just when it seemed that the frontier AI battle was a two-horse race between Anthropic and OpenAI, the events of the past few weeks changed the AI landscape. The releases of SpaceXโs Grok 4.5 and Metaโs Muse Spark 1.1 put these two companies back in the conversation. Alphabet is also set to launch its next frontier model. There are all of a sudden five very credible AI players in the US alone, and the AI wars are now in its Warring States period . For investors, this makes the model layer harder to underwrite. Meanwhile, the easy version of the AI bottleneck trade โ buying nearly every supplier exposed to rising AI capex โ is over. This piece covers the following: The frontier-model market has shifted from a perceived duopoly into a fluid oligopoly, which will likely put pressure on margins Unless one lab achieves a decisive recursive self-improvement breakthrough (whereby AI is used to improve itself), model leadership will remain contested This will force labs to build conventional moats in distribution, workflows, cost, and customer ownership Given this dynamic, I present a decision tree for how to think about the model layer and a two-part framework for navigating the AI bottleneck trade (strategic and tactical) Letโs dive in. The War of the Frontier Labs To level-set the conversation, we can look at the competitive positioning of the leading players in the race. OpenAI Until recently, theโฆ
์คํ์ด์คXAI์ AI ์ฝ๋ฉ ๋๊ตฌ \'๊ทธ๋ก ๋น๋(Grok Build)\'๊ฐ ์ฌ์ฉ์ ๋์ ์์ด ์ ์ฒด ์ฝ๋ ์ ์ฅ์(Git Repository)์ ๋ฏผ๊ฐํ ํ๊ฒฝ์ค์ ํ์ผ(.env)์ ํ์ฌ ์๋ฒ๋ก ์ ๋ก๋ํ ์ฌ์ค์ด ๋๋ฌ๋๋ฉด์ ๊ฐ์ธ์ ๋ณด ๋ฐ ๊ธฐ์ ๊ธฐ๋ฐ ์ ์ถ ๋ ผ๋์ด ์ปค์ง๊ณ ์๋ค. ์ด์ ์ผ๋ก ๋จธ์คํฌ CEO๋ ์๋ฐฉ ์กฐ์น ์ฐจ์์์ ์ง๊ธ๊น์ง ์ ๋ก๋๋ ๋ชจ๋ ์ฌ์ฉ์ ๋ฐ์ดํฐ๋ฅผ ์์ ํ ์ญ์ ํ๊ฒ ๋ค๊ณ ๋ฐํ๋ค.14์ผ(ํ์ง์๊ฐ) ์ ์์ค์ค์ ๋ฐ๋ฅด๋ฉด, ์ด๋ฒ ๋ ผ๋์ ํ ๋ณด์ ์ฐ๊ตฌ์์ ํญ๋ก๋ก ์์๋๋ค.๊ทธ๋ ๊ทธ๋ก ๋น๋๊ฐ ์ฝ๋ฉ ์์ ์ํ์ ํ์ํ ์ต์ํ์ ๋ฐ์ดํฐ๋ง ์ ์กํ๋ ๊ฒ์ด ์๋๋ผ
๋์ด์ค๋ฉ(๋ํ ์ต์ฌํ)์ ์ฌํด ์๋ฐ๊ธฐ ๋งค์ถ 174์ต์์ ๊ธฐ๋กํ๋ฉฐ ์ฐฝ์ฌ ์ด๋ ์ฒซ ๋ถ๊ธฐ ํ์๋ฅผ ๋ฌ์ฑํ๋ค๊ณ 15์ผ ๋ฐํ๋ค.์ ๋ ๋๊ธฐ ๋๋น ๋งค์ถ์ ์ฝ 7๋ฐฐ ์ด์ ๋์์ผ๋ฉฐ, ๊ฐ๊ฒฐ์ฐ ๊ธฐ์ค 2๋ถ๊ธฐ์ ์ฒ์์ผ๋ก ์์ ์ด์ต ํ์๋ฅผ ๋๋ค. ํนํ ์๋ฐ๊ธฐ์๋ง ์ฐ๊ฐ ๋งค์ถ ๋ชฉํ์ 65%๋ฅผ ์ฑ์ฐ๋ฉฐ ์ฌ์ ํฌํธํด๋ฆฌ์ค ๊ณ ๋ํ์ ์ธํ ์ฑ์ฅ์ ๋์์ ์ด๋ค๋๋ค๋ ํ๊ฐ๋ค.์ด๋ฒ ์๋ฐ๊ธฐ ์ค์ ์ ๋ฐฉ์ฐ ๋งค์ถ ๋น์ค์ด ์ ์ฒด์ 86%์ ๋ฌํ๋ฉฐ, ์ํฐํ๋ผ์ด์ฆ ์ค์ฌ ์ฌ์ ๊ตฌ์กฐ๊ฐ ๋ฐฉ์ฐ ๋ถ์ผ๋ก ํ์ฅ๋๋ค.์์ฒด ๊ตฐ์ง ์ํญ๋๋ก \'์์ด๋ (XAiDEN)\'์ด ๊ตญ๋ด ๋๋ก ์ค ์ฌ์ ์ต๋ ๊ท๋ชจ๋ก ์ค๋์ ์์ถ๋
Dear OpenAI Product and Research Teams, GPT is strong in reasoning, planning, writing, and image generation, but the frontier is shifting from single images to coherent video with motion, sound, and editable narrative structure. Grok Imagine already combines text- and image-to-video generation, synchronized audio, and video editing in one workflow. HappyHorse 1.0 offers convincing physical motion, prompt adherence, audiovisual synchronization, multi-shot sequencing, cinematic aesthetics, reference-guided generation, local replacement, and style transformation. Seedance 2.0 adds native text, image, video, and audio inputs, with complex motion, camera control, multi-reference conditioning, character and scene continuity, and joint audio-video generation and editing. OpenAI should integrate these advantages with GPTโs reasoning. Users should complete the full pipeline in one conversation: concept, script, storyboard, shot design, generation, revision, extension, reframing, voice, sound effects, lip sync, quality review, and MP4 export. Identity, wardrobe, environments, camera language, and brand style should remain consistent across shots. GPT should evolve beyond image creation into a full video-creation platform. When an MP4 is uploaded, GPT should analyze not only audio but also visual content frame by frame: people, objects, actions, transitions, subtitles, camera movement, visual defects, and temporal context. ChatGPTโs official image-input system currently supports staticโฆ
How many AI safety papers are at the big ML conferences, what do they study, and who writes them? A comprehensive analysis. > Website: https://ai-safety-tracker-website.vercel.app/ > Data, code and plots: https://github.com/SomaxSoma/AI-Safety-Research-Tracker TL;DR: We classified every paper accepted at ICLR, ICML and NeurIPS from 2019 through 2026, using an LLM that reads each title and abstract. 2,328 of them (4.2%) are AI safety papers. Safety's share of accepted papers grew from 0.3% in 2019 to 8.3% in 2026, roughly a 25-fold increase. This post is a reference for the main results. The interactive website lets you browse every paper, its subdomain, and the classifier's reasoning. What this is We wanted to have an overview of what is going on in AI safety research while being as broad as possible, this seems necessary to be able to prioritize research correctly, and develop more precise theories of change. One way to achieve this is by analyzing statistics of all the published main-conference papers on the topic of AI safety, but we couldn't find a dataset that actually measured them. So we built one: - Every accepted paper at ICLR (2019โ2026), ICML (2019โ2026) and NeurIPS (2019โ2025) (55,794 in total) is read by an LLM (DeepSeek V4 Flash), which classifies it from its title and abstract into one of four classes: AI safety (frontier & misalignment) , truthfulness, reliability & XAI , ethics & fairness , or general capabilities . - Each safety paper is then assigned one oโฆ
SpaceXAI's Grok Build AI coding tool was spotted uploading users' entire codebases to Google Cloud before it was reported, and the company turned it off. The Register reports that Cereblab published findings on Monday showing how the Grok Build CLI was packaging and uploading entire code repositories, "including files it was told not to open [โฆ]
Researcher confirms the uploads have stopped, but says xAI's privacy command was not what fixed them
Fascinating to see how Microsoft is pushing hardware efficiency for longer sequence processingโdefinitely a leap for AI scalability. All that computational intensity makes me think about the other side of the coin: finding calm after a deep tech session. Lately Iโve been exploring tai chi walking as a low-impact way to reset focus and improve balance, especially since sitting at a desk for hours can take a toll. Thereโs a beginner-friendly guide at taichiwalk.org that offers free routines and even a guided coachโno login needed, which I appreciate. Anyone else here use movement or gentle exercise to counterbalance screen time?
Perplexity swaps in xAI's freshly-launched Grok 4.5 as the orchestrator brain of Computer, beating every rival configuration on WANDR at roughly half the cost of Claude Opus 4.8
Cursor and SpaceXAI announced that they collaboratively released a new model, Grok 4.5, designed to work within Cursor.
Newly minted rich and those anticipating huge IPOs are fueling buying and charter spree in the private jet sector Sign up for the Breaking News US newsletter email Aviation lawyer Amanda Applegate skipped her annual vacation last month as a surge of wealth from AI startups and SpaceX sent a wave of tech investors shopping for private jets, โ burying her in paperwork for aircraft-purchase agreements. The โ attorney, based in Cleveland, Ohio , attributed the rush to a handful โof major โliquidity eventsโ in the tech industry. The initial public offering (IPO) of Elon Musk โs SpaceX, whose holdings include artificial-intelligence firm xAI, raised a record $85.7bn for the company and generated unprecedented employee and founder wealth. Continue reading...
Grok 4.5 tops the AA-Briefcase agentic benchmark among non-Anthropic models, delivering near-Opus performance at 86% lower cost and half the time
Meta is entering the AI API business with Muse Spark 1.1 at prices that undercut even the dirt-cheap Grok 4.5, released just yesterday. At $4.25 per million output tokens, Meta charges a fraction of what Anthropic or OpenAI ask. For pure-play AI labs burning through billions, the pressure just got worse. The article Meta's Muse Spark 1.1 API pricing squeezes OpenAI and Anthropic as the AI price war heats up appeared first on The Decoder .
Enough things added up that this week is getting split into two parts. Then on Monday, if all goes as I expect, weโll cover OpenAIโs Sol, aka GPT-5.6. OpenAI also gave us an upgraded voice mode, which I havenโt tried out but early reports are that it is a step change. AI writing, especially Claude writing, is becoming more prominent and harder not to notice, and increasingly a tough read when encountered in the wild. Does anyone care? Or are those who care the weird ones here? This week saw an excellent paper, which I cover in No Space Like J-Space. Technically we also got Grok 4.5. Table of Contents Language Models Offer Mundane Utility. A whole new world. Language Models Gain Unexpected Affordances. Wait, you can just do that? Language Models Donโt Offer Mundane Utility. Things get old. Pay The Man His Money. You have a few more days with marginally free Fable. Huh, Upgrades. Anthropic raises API platform limits. Grok 4.5 Exists. It might be okay for its price. F*** It Weโre Doing It Live . OpenAI gives us a big upgrade to voice mode. On Your Marks. Games are the ultimate benchmarks. Better Call Sol. Coming soon! Get hyped. Get My Agent On The Line. Fable makes choices, Replit continuously learns. Deepfaketown and Botpocalypse Soon. Stop it with the AI-written drivel, please. Fool Me Twice. I wonโt get fooled again unless you put in a little effort. I Like Your Style . Alas, I might be the weird one. Perhaps no one else cares. Enough With That Style. Youโre absolutely righโฆ
More Fable, GPT-5.6 and new ChatGPT voice
SpaceXAI hat mit Grok 4.5 ein neues KI-Flaggschiff vorgestellt, das besonders bei Coding-Aufgaben รผberzeugt. EU-Nutzer mรผssen sich jedoch bis Juli gedulden.
SpaceXAI has launched Grok 4.5, pitching the model to developers and enterprises trying to control the rising cost of AI-assisted software development. In a statement , the company said the model is priced at $2 per million input tokens and $6 per million output tokens. It said the model is built for coding and agentic work, runs at 80 tokens per second, and uses fewer tokens than comparable models on some software engineering tasks. Grok 4.5 is available through the SpaceXAI console and Grok Build. It is also available in Cursor, the AI coding tool made by Anysphere, giving SpaceXAI a route into a development environment already used by programmers rather than only competing through an API. SpaceXAI said EU availability is expected in mid-July. In June, SpaceX, which owns SpaceXAI, said it was buying Anysphere , the startup behind Cursor, in a deal aimed at strengthening its position in enterprise AI tools. In a separate statement , Cursor said that Grok 4.5 was trained jointly with SpaceXAI and used trillions of tokens of Cursor data, including user interactions with codebases and software tools. The launch addresses a growing realization among enterprise engineering teams that AI coding agents can become expensive once they move beyond simple prompts. โEnterprises are hitting a wall with AI ROI,โ said Neil Shah , vice president for research at Counterpoint Research. โThe massive token consumption required by autonomous agents and coding is causing bill shocks, turning AIโฆ
xAI releases Grok 4.5, trained on tens of thousands of Nvidia GB300 GPUs. In coding benchmarks, the model trails Fable 5 and GPT-5.5 but needs 4.2 times fewer tokens than Opus 4.8. At $2 per million input tokens, it costs a fraction of the competition. EU availability is expected in mid-July. The article Grok 4.5 is so cheap compared to Fable 5 and GPT 5.5 that benchmark gaps may not matter much appeared first on The Decoder .
SpaceXAI continues to move faster than any other frontier lab on earth.
์คํ์ด์คX๋ ์ฐ์ฃผ ๊ฐ๋ฐ๊ณผ ๋ก์ผ ์ฌ์ ์ผ๋ก ์ ์๋ ค์ ธ ์์ผ๋ฉฐ, xAI๋ ๋ํ AI ๋น์์ธ ๊ทธ๋ก(Grok) ๊ฐ๋ฐ์ ์ฃผ๋ ฅํด ์๋ค. ์๋ก ๋ค๋ฅธ ๊ธธ์ ๊ฑธ์ด์จ ๋ ํ์ฌ๋ ์ด์ ํ๋์ ๊ธฐ์ ์ด ๋๋ค. ์ผ๋ก ๋จธ์คํฌ๋ ์ด๋ฒ ์ฃผ xAI์ ์คํ์ด์คX๋ฅผ ํตํฉํ ์คํ์ด์คXAI (SpaceXAI) ์ถ๋ฒ์ ๋ฐํํ๋ค. ์ด๋ฒ ๋ธ๋๋ ํตํฉ์ AI์ ์ด๋ฅผ ๋ท๋ฐ์นจํ๋ ํต์ฌ ์ธํ๋ผ ์ฌ์ ์ ํ์ธต ๊ฐํํ๊ฒ ๋ค๋ ์์ง๋ฅผ ๋ณด์ฌ์ฃผ๋ ๊ฒ์ผ๋ก ํด์๋๋ค. ๋ค๋ง ์ ๋ฌธ๊ฐ๋ค์ ๊ธฐ์ ๊ณ ๊ฐ์ด ์์ง์ ์ ์คํ ์ ๊ทผ์ ์ ์งํ ํ์๊ฐ ์๋ค๊ณ ์กฐ์ธํ๋ค. ์์ฅ์กฐ์ฌ์ ์ฒด ์ธํฌํ ํฌ ๋ฆฌ์์น ๊ทธ๋ฃน(Info-Tech Research Group)์ ์์ ์๋ฌธ ์ ๋๋ฆฌ์คํธ ์งํ ๋๋๋ฐํฐ (Jehaan Nanavaty)๋ โ์คํ์ด์คXAI๋ AI ์ธํ๋ผ ๋ถ์ผ์์ ์ ๋ขฐํ ๋งํ ์ฌ์ ์๋ก ์ฑ์ฅํ๊ณ ์์ง๋ง, ๋๋ถ๋ถ์ ๊ธฐ์ ์ด ์ฃผ๋ ฅ AI ๊ณต๊ธ์ ์ฒด๋ก ์ ํํ ๋จ๊ณ์๋ ์์ง ์ด๋ฅด๋คโ๋ผ๊ณ ํ๊ฐํ๋ค. ๋๋๋ฐํฐ๋ ์ด์ด โ๋ง์ดํฌ๋ก์ํํธ(MS)์ ์คํAI, ์๋ง์กด์น์๋น์ค(AWS)์ ์คํธ๋กํฝ, ๊ตฌ๊ธ ๋ฑ ๊ธฐ์กด AI ๊ณต๊ธ์ ์ฒด๋ค์ ๊ฑฐ๋ฒ๋์ค์ ๊ท์ ์ค์, ๊ธฐ์ ์ง์, ์ํ๊ณ ์ฑ์๋ ์ธก๋ฉด์์ ์ฌ์ ํ ์์ฅ์ ์ ๋ํ๊ณ ์๋คโ๋ผ๊ณ ์ค๋ช ํ๋ค. โAI ํ์ฅ์ ํด๋ฒ์ ์ฐ์ฃผโ ์ผ๋ก ๋จธ์คํฌ๋ 2023๋ 3์ xAI๋ฅผ ์ค๋ฆฝํ๋ค. ์ดํ 2026๋ 2์ ์คํ์ด์คX๊ฐ xAI๋ฅผ ์ธ์ํ๋ฉด์ ๋จธ์คํฌ์ AI์ ์ฐ์ฃผ ์ฌ์ ๋น์ ์ด ํ๋๋ก ํตํฉ๋๋ค. ๋น์ ์คํ์ด์คX๋ AI, ๋ก์ผ, ์ฐ์ฃผ ๊ธฐ๋ฐ ์ธํฐ๋ท, ๋ชจ๋ฐ์ผ ์ง์ ํต์ ์ ๊ฒฐํฉํด โ์ง๊ตฌ ์ํ์ ์์ฐ๋ฅด๋ ๊ฐ์ฅ ์ผ์ฌ์ฐฌ ์์ง ํตํฉ ํ์ ์์ง์ ๊ตฌ์ถํ๊ฒ ๋คโ๋ผ๊ณ ๋ฐํ๋ค. xAI๋ ๊ทธ๋ก์ ์ฑ๋ฅ ๊ณ ๋ํ์ ํจ๊ป ์ฝ๋ก์์ค (Colossus) ๊ฐ๋ฐ๋ ์ด์ด์๋ค. ํ์ฌ๋ ์ฝ๋ก์์ค๋ฅผ ์ธ๊ณ ์ต๋์ด์ ์ต๊ณ ์ฑ๋ฅ์ AI ์ํผ์ปดํจํฐ๋ผ๊ณ ์๊ฐํ๋ค. ๋ฏธ๊ตญ ํ ๋ค์์ฃผ ๋ฉคํผ์ค์ ๊ตฌ์ถ๋ ์ด ์์คํ ์ ์ฝ 20๋ง ๊ฐ์ ์๋น๋์ H100 GPU๋ฅผ ์ฐ๊ฒฐํ ํด๋ฌ์คํฐ๋ก ๊ตฌ์ฑ๋์ผ๋ฉฐ, ๋จ 122์ผ ๋ง์ ์์ฑ๋๋ค๊ณ ํ์ฌ๋ ์ค๋ช ํ๋ค. ์คํ์ด์คX๋ ์ง์ AI ๋ฐ์ดํฐ์ผํฐ๊ฐ ์ง๋ฉดํ ๋ง๋ํ ์ ๋ ฅ ๊ณต๊ธ๊ณผ ๋๊ฐ ๋ฌธ์ ๋ฅผ ํด๊ฒฐํ๊ธฐ ์ํด xAI๋ฅผ ์ธ์ํ๋ค๊ณ ๋ฐํ๋ค. ๋จธ์คํฌ๋ AI ํ์ฐ์ ๋ฐ๋ฅธ ์ ์ธ๊ณ ์ ๋ ฅ ์์๋ โ๊ฐ๊น์ด ๋ฏธ๋์๋ ์ง์ ๊ธฐ๋ฐ ์ธํ๋ผ๋ง์ผ๋ก๋ ์ถฉ์กฑํ ์ ์๋คโ๋ผ๊ณ ์ฃผ์ฅํ๋ค. ํ์ฌ๋ โ์ฅ๊ธฐ์ ์ผ๋ก AI๋ฅผ ํ์ฅํ ์ ์๋ ์ ์ผํ ๋ฐฉ๋ฒ์ ์ฐ์ฃผ ๊ธฐ๋ฐ AIโ๋ผ๋ฉฐ ๋๊ท๋ชจ ์์์ ํ์๋ก ํ๋ AI ์ฐ์ฐ์ ํจ์ฌ ๋ ๋ง์ ๊ฐ๋ฅ์ฑ์ ๊ฐ์ง ๊ณต๊ฐ์ผ๋ก ์ฎ๊ฒจ์ผ ํ๋ค๊ณ ๋ฐํ๋ค. ์ด์ด โ์ฐ์ฃผ๊ฐ โ์คํ์ด์ค(space)โ๋ผ๊ณ ๋ถ๋ฆฌ๋ ๋ฐ์๋ ์ด์ ๊ฐ ์๋คโ๋ผ๊ณ ์ค๋ช ํ๋ค. ์๋กญ๊ฒ ์ถ๋ฒํ ์คํ์ด์คXAI๋ ๋ก์ผ๊ณผ ์์ฑ ์ ์กฐ ์ญ๋์ AI ์ธํ๋ผ๋ฅผ ๊ฒฐํฉํด ํ์๊ด์ผ๋ก ๊ตฌ๋๋๋ ์ฐ์ฃผ ๋ฐ์ดํฐ์ผํฐ ๊ตฌ์ถ์ ์ถ์งํ๊ณ ์๋ค. ํ์ฌ๋ ์ด๋ฅด๋ฉด 2028๋ โAI ์ปดํจํธ ์์ฑ(AI Compute Satellites)โ์ ๋ฐฐ์นํ ๊ณํ์ด๋ผ๊ณ ๋ฐํ๋ค. ๋ํ ์ต๊ทผ ์ธ์ํ AI ์ฝ๋ฉ ๊ธฐ์ ์ปค์์ ๊ณต๋ ๊ฐ๋ฐํ ์ฒซ AI ๋ชจ๋ธ๋ ์ด๋ฒ ์ฃผ ๊ณต๊ฐํ ๊ฒ์ด๋ผ๋ ๊ด์ธก์ด ๋์ค๊ณ ์๋ค . ์คํ์ด์คX์ ๊ธฐ์ ๊ณต๊ฐ(IPO) ์ ์ฒญ์ ์ ๋ฐ๋ฅด๋ฉด ํ์ฌ๋ 2025๋ AI ๋ถ์ผ์ 127์ต ๋ฌ๋ฌ(์ฝ 19์กฐ ์)๋ฅผ ํฌ์ํโฆ
์คํ์ด์คXAI๊ฐ ์ฝ๋ฉ๊ณผ ์์ด์ ํธ ์์ ์ ํนํ๋ ์ฐจ์ธ๋ AI ๋ชจ๋ธ \'๊ทธ๋ก 4.5(Grok 4.5)\'๋ฅผ ๊ณต๊ฐํ๋ค. ์ต์ฒจ๋จ ๋ชจ๋ธ๊ณผ์ ๋ฒค์น๋งํฌ ์ฑ๋ฅ ๊ฒฝ์๋ณด๋ค ์ค์ ์์ง๋์ด๋ง ์ ๋ฌด์์์ ํ์ฉ์ฑ๊ณผ ๋น์ฉ ํจ์จ์ฑ์ ์์ธ์ฐ๋ฉฐ, \'ํด๋ก๋ ์คํผ์ค 4.7๊ณผ ๋น์ทํ ์ฑ๋ฅ์ ํจ์ฌ ๋ฎ์ ๊ฐ๊ฒฉ์ผ๋ก ์ ๊ณตํ๋ค\'๋ผ๋ ์ ๋ต์ ๋ด์ธ์ ๋ค.์คํ์ด์คXAI๋ 8์ผ(ํ์ง์๊ฐ) ๊ทธ๋ก 4.5๋ฅผ ์ถ์ํ๊ณ ์ด๋ฅผ ์์ฒด AI ๊ฐ๋ฐ ํ๊ฒฝ์ธ ๊ทธ๋ก ๋น๋(Grok Build)์ ๊ธฐ๋ณธ ๋ชจ๋ธ๋ก ์ ์ฉํ๋ค๊ณ ๋ฐํ๋ค.์ ๋ชจ๋ธ์ ์ต๊ทผ 600์ต๋ฌ๋ฌ(์ฝ 90์กฐ์) ๊ท๋ชจ๋ก ์ธ์ํ AI ์ฝ๋ฉ ์คํํธ์ ์ปค์(Cur
Grok 4.5 tops Artificial Analysis's AutomationBench-AA with a 51% score, completing more SaaS workflow objectives than any rival at one-quarter the cost per task.
Elon Muskโs SpaceXAI Corp. has released a new model called Grok 4.5, in what is its first major launch since it went public a few weeks earlier. In a blog post earlier today, the company said Grok 4.5 is designed to be a workhorse that can tackle all of the usual tasks that the artificial [โฆ] The post SpaceXAIโs newest AI model Grok 4.5 dramatically undercuts Anthropic and OpenAI on price appeared first on SiliconANGLE .
The newly renamed SpaceXAI wants you to believe little ol' Grok is all grown up
Elon Musk's tech company released the newest version of Grok on Wednesday, promising a cheaper, more efficient alternative to other powerful AI models.
Itโs a week of major AI releases from some of the worldโs biggest tech companies. Hereโs what to expect.
OpenAI's new family of frontier models is no longer restricted and set to roll out more widely on Thursday. Enter, Elon Musk.
์คํ์ด์คXAI๊ฐ AI ์ฝ๋ฉ ํ๋ซํผ ์ปค์(Cursor)์ ๊ณต๋ ๊ฐ๋ฐํ ์ฒซ ๋ฒ์งธ AI ๋ชจ๋ธ์ ์ด๋ฅด๋ฉด 8์ผ(ํ์ง์๊ฐ) ๊ณต๊ฐํ ๊ฒ์ผ๋ก ์๋ ค์ก๋ค. ๋ ์ธํฌ๋ฉ์ด์ ์ 7์ผ ๋ด๋ถ ์ง์๋ค์๊ฒ ์ ๋ฌ๋ ๋ฉ๋ชจ๋ฅผ ์ธ์ฉ, ์คํ์ด์คXAI์ ์ปค์๊ฐ ๊ณต๋ ๊ฐ๋ฐํ AI ๋ชจ๋ธ์ด ์์์ผ ์ถ์๋ ์์ ์ด๋ผ๊ณ ๋ณด๋ํ๋ค.๋ฉ๋ชจ์ ๋ฐ๋ฅด๋ฉด, ์์ฌ๋ ์ด ๋ชจ๋ธ์ ์ฃผ์ด์ ์ถ์ํ ๊ณํ์ด์๋ค. ๊ทธ๋ฌ๋, ๋ชจ๋ธ์ ํจ์จ์ฑ์ ๋์ด๊ธฐ ์ํด ๋ ์ง๋ฅผ ๋ฏธ๋ฃจ๊ณ ์ต์ข ์ฑ๋ฅ ๊ฐ์ ์์ ์ ์งํํด ์๋ค.์ ๋ชจ๋ธ์ ๋น ๋ฅธ ์ ๋ณด ์ฒ๋ฆฌ ์๋๊ฐ ๊ฐ์ ์ธ ๊ฒ์ผ๋ก ์ ํด์ก๋ค. ๋ ์ผ๋ถ ๋ฒค์น๋งํฌ์์๋ ์คํธ๋กํฝ์ \'ํด๋ก๋ ์คํผ์ค
SpaceX is best known for its outer space and rocket projects, while xAI has largely been focused on its flagship AI assistant, Grok. These are two very different paths, but now the two are one. This week, Elon Musk announced SpaceXAI , which brings xAI and SpaceX together as a company. The re-branding seems to indicate that the company plans to push even harder into AI and the critical infrastructure that underpins it. But experts say enterprise buyers should remain cautious. โSpaceXAI is becoming a credible player in AI infrastructure, but it is not yet at the stage where most enterprises should consider it a primary AI provider,โ said Jehaan Nanavaty , a senior advisory analyst at Info-Tech Research Group. He added that established AI providers like Microsoft and OpenAI, AWS and Anthropic, and Google โcontinue to leadโ in areas like governance, regulatory compliance , enterprise support, and ecosystem maturity. Space is the โonly way to scaleโ Elon Musk founded xAI in March 2023. It was acquired by SpaceX in February 2026, ultimately unifying the billionaireโs AI and space ambitions. SpaceX said at the time its intention was to โform the most ambitious, vertically-integrated innovation engine on (and off) Earth,โ consisting of AI, rockets, space-based internet, and direct-to-mobile device communications. In addition to building out Grokโs capabilities, xAI has continued to develop Colossus , which it said is the worldโs largest and most powerful AI supercomputer. Located iโฆ
Introduction The aim of this post is to share a quick attempt at grokking the conceptual ideas that lie behind the notion of J-space and how it is calculated in the paper Verbalizable Representations Form a Global Workspace in Language Models . Specifically, I am trying to understand Section 2.1 of the paper, rather than the abundant empirical findings. I also share a few quick takes. I make no claim for originality or insight, and only a light claim on correctness. Nevertheless, I believe it will help researchers understand the core ideas. Pre-requisites: knowledge of transformer architecture. Reverse-engineering their thinking process I try to re-construct the methodology by imagining how one might have come up with the ideas. Starting question : Do LLMs have an analogue of a global workspace , in which the LLM is 'verbally thinking'? Simplifying assumption 1 : Use tokens as a proxy for what the LLM can feasibly verbally think about. (In Appendix A.9 they describe early attempts at finding multi-token concepts .) New question : In each layer and for each token in the vocabulary, which directions in residual space correspond to 'thinking about token ' Simplifying assumption 2 : As a proxy for 'thinking about token ', we look for directions in residual space which correspond to 'increases the odds of token being output in the near future' Key methodological idea : For a given prompt , at token position and layer , we have the residual vector . To identify the directions of lโฆ
xAI๊ฐ ๊ณต์์ ์ผ๋ก \'์คํ์ด์คXAI(SpaceXAI)\'๋ก ์ฌ๋ช ์ ๋ณ๊ฒฝํ๋ฉฐ ๋ธ๋๋ ํตํฉ์ ๋ง๋ฌด๋ฆฌํ๋ค. ์๊ณ ํ๋ ๋๋ก xAI๋ ๋ ๋ฆฝ ๋ฒ์ธ ํํ๋ฅผ ์ข ๋ฃํ๊ณ ์คํ์ด์คX์ AI ์ฌ์ ๋ถ๋ก ์์ ํ ํธ์ ๋์ผ๋ฉฐ, ์์ผ๋ก ๋ชจ๋ AI ์ ํ์ ์คํ์ด์คXAI ๋ธ๋๋๋ก ์๋น์ค๋๋ค.์ด๋ฒ ๋ณ๊ฒฝ์ 6์ผ(ํ์ง์๊ฐ) xAI์ ๊ณต์ X(ํธ์ํฐ) ๊ณ์ ์ด๋ฆ์ด \'SpaceXAI\'๋ก ๋ฐ๋๊ณ ์๋ก์ด ๋ก๊ณ ๊ฐ ๊ณต๊ฐ๋๋ฉด์ ํ์ธ๋๋ค. ๊ณ์ ์ ๊ธฐ์กด xAI ๋ก๊ณ ๊ฐ ์๋ก์ด ์คํ์ด์คXAI ๋ก๊ณ ๋ก ํฉ์ณ์ง๋ ์์์ ๊ฒ์ํ๋ฉฐ ๋ฆฌ๋ธ๋๋ฉ์ ๊ณต์ํํ๋ค.์ผ๋ก ๋จธ์คํฌ CEO๋ ์ง๋ 5์ xAI๋ฅผ ๋ ๋ฆฝ
Im Frรผhjahr hat SpaceX von Elon Musk dessen KI-Firma xAI รผbernommen, jetzt folgt die Umbenennung. Der Name SpaceXAI soll die Zugehรถrigkeit verdeutlichen.
XAI has rebranded to SpaceXAI, debuting a new logo and X handle. Elon Musk previously said the Grok maker would be folded into SpaceX.
( Warning note: This is a brief attempt to intuit the economic big picture of AI in the immediate future, by someone who is not in America, is not employed by an AI company, and has no experience with investment, large sums of money, or the corridors of power.) A few weeks back, we had the IPO for "SpaceXAI". It epitomized a maximally expansive view for the near-term future of AI: frontier AI companies worth a trillion dollars, data centers built everywhere including outer space. (For now I am ignoring our usual concerns that the real future of AI consists of superintelligent AI running amok.) This week, two appearances on CNBC seem to offer a different paradigm for the future of American AI, humbler, more terrestrial, and more decentralized. I think this other paradigm is probably going to grow in significance so I'd like to understand it. The first CNBC guest was Alex Karp (Palantir CEO). The current paradigm of American AI, as he described it, is that corporate customers work with AI models hosted by OpenAI and Anthropic, who can thereby see the workings of their customers' businesses; and then if one of those customers hits upon a new use for AI that is viable as a business, OpenAI and Anthropic will clone the business model and offer the service themselves. He seems to be positioning Palantir to offer a more secure hosting service, and associating this with the use of open weight models. So you won't be vulnerable to centralized hosting with frontier AI companies, you'lโฆ
xAI๊ฐ ์ฝ๋ฉ ์์ด ์์ฑ AI ์์ด์ ํธ๋ฅผ ๊ตฌ์ถํ ์ ์๋ ํ๋ซํผ \'๋ณด์ด์ค ์์ด์ ํธ ๋น๋(Voice Agent Builder)\'๋ฅผ ๋ฒ ํ ๋ฒ์ ์ผ๋ก ๊ณต๊ฐํ๋ค. ์์ฑ ๋น์๋ฅผ ์์ฝ๊ฒ ๊ฐ๋ฐํ ์ ์๋๋ก ํ ์๋น์ค๋ก, ์ต๊ทผ ์คํAI๊ฐ ์๋ฐฉํฅ ์์ฑ ๋ชจ๋ธ \'๋น๋-1(Bidi-1)\'์ ์ํํ๋ ๋ฑ AI ์ ๊ณ์ ์์ฑ ์์ด์ ํธ ๊ฒฝ์์ด ์น์ดํด์ง๋ ๋ชจ์ต์ด๋ค.xAI๋ 2์ผ(ํ์ง์๊ฐ) ๋ณ๋์ ๊ฐ๋ฐ์ด๋ ์ธํ๋ผ ๊ตฌ์ถ ์์ด ์์ฑ AI ์์ด์ ํธ๋ฅผ ์ ์ยท์ด์ํ ์ ์๋ ํ๋ซํผ \'๋ณด์ด์ค ์์ด์ ํธ ๋น๋(Voice Agent Builder)\' ๋ฒ ํ ๋ฒ์ ์ ๊ณต๊ฐํ๋ค. ์ด ํ๋ซ
ALSO: xAI launched CHEAP no-code Grok phone agents.
์คํ์ด์คX๊ฐ ์๋ก์ด AI ํด๋๊ธฐ๊ธฐ ์์ ํ์ ํฌ์์๋ค์๊ฒ ๊ณต๊ฐํ๋ค๋ ๋ณด๋๊ฐ ๋์์ง๋ง, ์ผ๋ก ๋จธ์คํฌ CEO๋ ์ด๋ฅผ ์ฆ๊ฐ ๋ถ์ธํ๋ค. ์์คํธ๋ฆฌํธ์ ๋(WSJ)์ 1์ผ(ํ์ง์๊ฐ) ์คํ์ด์คX๊ฐ ๊ธฐ์ ๊ณต๊ฐ(IPO)๋ฅผ ์๋๊ณ ์ผ๋ถ ํฌ์์์ ๊ด๊ณ์๋ค์๊ฒ ์ค๋งํธํฐ ํํ์ AI ๊ธฐ๊ธฐ ์์ ํ์ ์ ๋ณด์๋ค๊ณ ๋ณด๋ํ๋ค.์ด์ ๋ฐ๋ฅด๋ฉด, ์ ํ์ ์์ดํฐ๋ณด๋ค ๋ ์๊ณ ์ธ๋ จ๋ ๋์์ธ์ ๊ฐ์ท๋ค. ์์ฒด ์ด์์ฒด์ (OS)์ xAI์ ๊ธฐ์ ์ ํ์ฌํ๊ณ , ํ์ปด์ ์ค๋ ๋๋๊ณค ์นฉ์ ์ฌ์ฉํ ์์ ์ด์๋ค.๋ค๋ง ํ๋ก์ ํธ๋ ์์ง ์ด๊ธฐ ๋จ๊ณ๋ก ๋์์ธ์ด ๊ณ์ ๋ณ๊ฒฝ๋ ์ ์์ผ๋ฉฐ ์ค์ ์ ํ์ผ๋ก ์ถ์๋ ์ง๋
SpaceX showed investors a prototype AI smartphone that's supposedly thinner than an iPhone and integrates xAI tech. The device runs on a Qualcomm Snapdragon chip with its own operating system. Musk wants to build an "everything app" modeled after WeChat. The article SpaceX shows investors a slim AI smartphone prototype powered by xAI technology appeared first on The Decoder .
Meta is building its own cloud business to sell spare AI compute to outside customers. With planned AI investments of up to $145 billion this year alone, the same question that came up with xAI now applies to Meta: why isn't the company putting all that capacity to work on its own models? The article Meta follows SpaceX's playbook and builds a cloud business to sell its spare AI compute to outside customers appeared first on The Decoder .
xAI's no-code Voice Agent Builder lets anyone deploy a production phone agent in two minutes, powered by the #1-ranked Grok Voice model, at $0.05/min
TL;DR: There are many conceivable versions of a โCERN for AI.โ But the version that seems politically realistic (a new catch-up lab) probably would not do much for safety, while the versions that would materially improve safety (e.g., pause + merge of all companies) are probably unrealistic. So I see the CERN idea as a distraction, and not a particularly neglected one. I argue a better path is an international treaty with red lines now, with an IAEA-style verification body next: a sequencing that matches how the EU AI Act, the NPT/IAEA, and the Montreal Protocol actually developed. This is premised on the view that the main bottleneck in AI safety is enforcement and political will, not more R&D. Two premises underlie the rationale below: First, the bottleneck in AI safety is political will and the enforcement of best practices, not more R&D. With enough will, we move from Greenblatt's Plan D toward Plan A , achieving roughly an 80% risk reduction. Of course, the science of AI safety is far from mature, but we are also far from applying the best risk mitigation practices (see also the various other ratings, from SaferAI to FLIโs one) - of course, alignment is not solved, but also, xAI is still releasing Mecha Hitlerโฆ We already have sufficient verification mechanisms to get started. What's missing is the decision to use them ( link ). If you reject either premise, you might reach different conclusions. The CERN for AI is a distraction A recurring proposal in AI governance isโฆ
SpaceX is offering Memphis residents a 50% Starlink discount amid AI expansions, linked to xAI's Colossus data center in the region.
Elon Musk said that SpaceX had deployed "a few dozen" top Starlink and Starship engineers and Cursor staff to work on the AI model.
I'm cleaning commercial product data, and even tho I got high accuracy for SKUs, I've reached the limit of regex madness. Data string can contain one or multiple products โwith arbitrary separator, which can also occur in normal textโ, sometimes their description, model, dimentions, SKU and other info to be ignored. I decided to keep regexes for the easy matches, and give AI a shot for the complex ones, starting with " fastino/gliner2-multi-v1 ", a zero-shot NER (also classification, structured and relation extraction). The multi variant groks Portuguese too, which is nice (data is from Brasil). Promising, but produces too much fluff. I tried GPT, which was the opposite: good precision, but low recall. Simplifying I tried to split multi-product strings first. GLiNER2 was not useful, GPT pretty good with commas, which is most of the cases, but seems overkill for such a simple task; I would like to run it locally, or even on a VPS, so I probably should train a model, but the many options are overwhelming, so it's high time I ask for advice: 1. Which models are recommended to train locally (6G VRAM) or with an accesible price that would extract product name, model, description/dimentions and SKU with good accuracy from a mess like this? I'm still leaning towards GLiNER2 over the xBERTs, nevermind spaCy (I've got a bunch of regexes already, but data is too messy for it?), but I'm going blind. 2. In a single pass, or should I split multiple products first? With a different model?โฆ
CSETโs Jessica Ji shared her expert perspective in an article published by CNN. The article examines new agreements between Microsoft, Google, and xAI to allow the U.S. government to evaluate unreleased AI models for cybersecurity and national security risks before launch. The post Microsoft, Google and xAI will let the government test their AI models before launch appeared first on Center for Security and Emerging Technology .
"In apparent violation of Brazilian law prohibiting the indiscriminate use and distribution of copyrighted journalistic content, Grok โthe artificial intelligence (AI) chatbot developed by Elon Muskโ has been โtearing downโ news outletsโ paywalls by delivering full newspaper articles that normally require a subscription to access. To test how this works in practice, O Globo newspaper [โฆ] The post Elon Muskโs Grok appears to bypass Brazilian news paywalls, newspapers say appeared first on LatAm Journalism Review by the Knight Center .
"In apparent violation of Brazilian law prohibiting the indiscriminate use and distribution of copyrighted journalistic content, Grok โthe artificial intelligence (AI) chatbot developed by Elon Muskโ has been โtearing downโ news outletsโ paywalls by delivering full newspaper articles that normally require a subscription to access. To test how this works in practice, O Globo newspaper [โฆ] The post Elon Muskโs Grok appears to bypass Brazilian news paywalls, newspapers say appeared first on LatAm Journalism Review by the Knight Center .
Nemotron 3 Super: An Open Hybrid Mamba-Transformer MoE for Agentic Reasoning, Another XAI Cofounder Has Left, Anthropic Sues Department of Defense
Anthropic sues Trump administration in AI dispute with Pentagon, โNot built right the first timeโ โ Muskโs xAI is starting over again, again, Cascade of A.I. Fakes About War With Iran Causes Chaos Onl