AI/ML News & Innovations Hub

AI/ML news, top picks, and generated innovation digests.

★ Visit ai-karthik.com
422Sources
35696News Items
8Top Picks
210Blogs
failedLast Run

Latest AI/ML News

35696 matching items

Persistent Truncation Issues with GPT-4o-Transcribe – Has Anyone Fully Solved This?
OpenAI Community 2026-08-13 00:18 UTC Score 43.0 AI-116-20260813-social-media-68394c13 Full article

Persistent Truncation Issues with GPT-4o-Transcribe – Has Anyone Fully Solved This?

hello, extended dev of own api here. i ve found that your solution is dramatically enhance quality, despite there is still some truncation b thankyou and openai is pointing in a completely wrong use of prompt too. because check what gemini says: to use the prompt like adjusting the he words with the pre vious transcrbe words. and that is completely not working.

OpenAI Community 2026-08-13 00:12 UTC Score 43.0 AI-116-20260813-social-media-a23bf5c1 Full article

Feature Request: Enforce Persistent Brand Rules Across Image Generation, Assets, QA, and Delivery

I’ve been working on this issue for the past few days, and I believe I have found a solution that is not documented in OpenAI’s documentation (at least I haven’t found it anywhere). Problem 1: ChatGPT doesn’t know much about itself. OpenAI’s documentation isn’t natively part of ChatGPT. Without searching OpenAI documentation, ChatGPT cannot tell you whether or not it can or can’t do something. Unfortunately, many times this means that ChatGPT will generate a workflow it reasonably believes it has the ability to achieve the desired outcome. Better put, ChatGPT asks ‘Can an AI model do this?’ and does a web search instead of ‘Can ChatGPT do this?’ and searches the OpenAI documentation. This problem will surface again, below. Problem 2: ChatGPT cannot locate assets in the Library. This is because of a huge flaw. Let’s say you have an image of your logo that is part of your Brand Bible, and that image is uploaded to your Library as “MyLogo.png.” When you upload the image, it is assigned a File ID, and that ID is used as the file name. So, “MyLogo.png” becomes “File_74829743974839.jpg,” for example. Issue 1: If your Brand Bible has a notation to use “MyLogo.png” as the canonical reference image, it has been renamed and cannot be found in your Library. Issue 2: The file type has changed. Issue 3. Even if you did know what the File ID was, and it was properly noted in the Brand Bible as the canonical reference, it still cannot be found in your Library. Enter Problem 3. Problem 3. C…

Feature suggestion: Image Generation Add-ons
OpenAI Community 2026-08-13 00:12 UTC Score 37.0 AI-116-20260813-social-media-eaaff4e5 Full article

Feature suggestion: Image Generation Add-ons

I would like to make a suggestion. Feature suggestion: Image Generation Add-ons I’m a ChatGPT Plus subscriber who uses image generation heavily for character artwork and creative projects. I regularly reach my image-generation limit, but upgrading to Pro is too expensive for me because I mainly need increased image generation rather than all of the additional Pro features. I’d love an optional image-generation add-on for Plus, with several levels—for example Image Boost 1–5. Higher levels could provide more daily image generations and faster generation speeds. Another option could be purchasable image-generation packs for users who temporarily need more generations. This would let customers pay specifically for additional image capacity without needing to upgrade their entire ChatGPT plan. I’d personally be willing to pay around 500 THB/month for a substantial image-generation upgrade. I think this could be especially useful for artists, designers, creators, and people using ChatGPT heavily for visual projects.

AI threatens nearly a quarter of Southeast Asia’s workforce but not all is lost
South China Morning Post AI 2026-08-13 00:00 UTC Score 42.0 AI-156-20260813-regional-ai--059be275 Full article

AI threatens nearly a quarter of Southeast Asia’s workforce but not all is lost

Fears of a new industrial revolution in which artificial intelligence replaces manual labour have yet to materialise in Southeast Asia – but white-collar workers should brace themselves. Nearly one in four workers in the region face generative AI (GenAI) disrupting or affecting their jobs, according to a July report by the International Labour Organization (ILO). Clerical, administrative and professional positions were most at risk, the report stated. Manual trades, craft and agricultural work...

Qdrant Blog 2026-08-13 00:00 UTC Score 60.0 USR-0074-20260813-ai-specialis-e0c70b84 Full article

Qdrant and Minima Deliver 2.92x More Agentic RAG Tasks per GPU-Hour

Reducing Retrieval and Calls When a retrieval-augmented generation (RAG) agent runs, it often has to plan a search, check the evidence it gets back, and try again when that evidence falls short. Those inefficiencies compound. Every extra retrieval and every extra model call adds latency, context, and inference cost.

Qdrant Blog 2026-08-13 00:00 UTC Score 37.0 USR-0074-20260813-ai-specialis-a5f9a2ff Full article

How Bayer Built an Enterprise-Scale Search Engine with Qdrant

Bayer is a global life sciences company operating at the intersection of two of the most consequential fields in human life: health and nutrition. Its pharmaceutical work supports drug discovery and patient care, while its crop science work supports food production at planetary scale. The company’s guiding ambition, “Health for all, hunger for none,” frames how it thinks about technology: AI is not a side project, but a lever applied across the entire organization, from improving the productivity of colleagues to accelerating yield prediction and drug discovery.

Apple Machine Learning Research 2026-08-13 00:00 UTC Score 54.0 AI-059-20260813-official-ai--6957dd9c Full article

When Unlearning Is Free: Leveraging Low Influence Points to Reduce Computational Costs

As concerns around data privacy in machine learning grow, the ability to unlearn, or remove, specific data points from trained models becomes increasingly important. While state of the art unlearning methods have emerged in response, they typically treat all points in the forget set equally. In this work, we challenge this approach by asking whether points that have a negligible impact on the model’s learning need to be removed. Through a comparative analysis of influence functions across language and vision tasks, we identify subsets of training data with negligible impact on model outputs…

Simon Willison Weblog 2026-08-12 23:59 UTC Score 69.0 USR-0110-20260812-ai-specialis-38e3dd60 Full article

DeepSeek V4 Pro 0813 (on OpenRouter)

DeepSeek V4 Pro 0813 (on OpenRouter) The latest DeepSeek Pro model is now available, via API only. I had to link to OpenRouter because DeepSeek don't have any obvious announcement page for their new model. I haven't been able to confirm if they plan to release the open weights, but given the weights are available for both April's deepseek-ai/DeepSeek-V4-Pro and July's deepseek-ai/DeepSeek-V4-Flash-0731 it seems likely. Update : the weights are now available on Hugging Face, 1.7T parameters, 893 GB. Interestingly I got very different looking pelicans for the three different reasoning levels of low, medium, and high. I've not noticed this kind of difference from any other model: Low: Medium: High: In terms of benchmarks... as far as I can tell those were released to the Official DeepSeek WeChat Group, then copied and pasted into a post on Reddit which was deleted by the moderators for being "low-effort", then copied into this ASCII-art table on Hacker News . Tags: ai , generative-ai , llms , pelican-riding-a-bicycle , deepseek , llm-release , ai-in-china

Nebius shares jump 34% on continued AI infrastructure demand
SiliconANGLE AI 2026-08-12 23:57 UTC Score 33.0 USR-0127-20260812-global-ai-ne-5fa97ec2 Full article

Nebius shares jump 34% on continued AI infrastructure demand

Shares of Nebius Group NV closed 34% higher today after it reported second-quarter earnings that topped expectations across the board. The Netherlands-based company operates a cloud platform optimized for artificial intelligence workloads. It also has two business units called Avride and TripleTen that offer autonomous driving software and programming courses, respectively. Nebius’ revenue surged 454% […] The post Nebius shares jump 34% on continued AI infrastructure demand appeared first on SiliconANGLE .

OpenAI Community 2026-08-12 23:46 UTC Score 45.0 AI-116-20260812-social-media-614b2dcc Full article

Kruel.ai KV2.0 - KX (experimental research) to current 8.2- Api companion co-pilot system with full modality , understanding with persistent memory

Pretty excited. A potential Future outcome signed an NDA with Kruel.Ai Inc. recently, and they will get access to the full end to end workings and math model to look at. Its been under a trial for about 3 months now. More excited to see what this Tech company thinks about it, not Everyday someone gets to learn how the Blackbox thinks under the hood of a system that is not like the others Lets see where this Adventure goes. updates have slowed down a little because there has been a lot of things going on with this around around me that has had me pretty nailed down. Meeting all the Company Partners next week This should be fun.

San Francisco-area estate sells for $70m in sign of AI-fueled wealth explosion
The Guardian AI 2026-08-12 23:22 UTC Score 45.0 AI-021-20260812-global-ai-ne-93d4bb2e Full article

San Francisco-area estate sells for $70m in sign of AI-fueled wealth explosion

New owner is said to be involved in AI industry, as new class of multimillionaires is entering the real estate race A luxury Lake Como-inspired estate in the San Francisco Bay Area has sold to a buyer “in the AI world” for a record $70m, in the latest sign of a new tech wealth boom that’s pushing home prices to extraordinary heights in the region The property south of San Francisco not only becomes the most expensive sale in its affluent town of Hillsborough, doubling a previous record, but is also the most expensive sale in northern California this year, according to listing agent Jennifer Gilson of Golden Gate Sotheby’s International Realty. Continue reading...

OpenAI Community 2026-08-12 23:11 UTC Score 37.0 AI-116-20260812-social-media-809417f6 Full article

Feature request: Bulk organization of existing chats into Projects

Hey @ meganyoungmee ! You make a good point about the difference between organizing new chats and cleaning up a backlog that may already contain hundreds of conversations. Moving chats into Projects one at a time works for occasional organization, but it becomes much less practical when someone wants to restructure a large existing history. Multi select, filtering, and bulk actions would make that kind of cleanup far more manageable. The approval step in your idea is important too. Having ChatGPT suggest where chats belong, then letting the user review those suggestions before anything moves, could make automation useful without taking control away from the user. I will make sure this feedback is shared internally. I cannot promise whether or when these controls might be added. - Sunny

Memory Limit Meter Display
OpenAI Community 2026-08-12 23:09 UTC Score 32.0 AI-116-20260812-social-media-825dd68f Full article

Memory Limit Meter Display

I’m going to second this and give it a bump. This is from 2025 and seems to have gone ignored. I’ll ask chat specifically about whether it’s getting close to it’s limit and it has to guess but can’t give an exact % of how much is left or used. It would be nice to know how close you are to hitting that limit cap. I was going to make a post about this but searched and found this one. Thank you

From the mine to the grid
New Scientist AI 2026-08-12 23:01 UTC Score 38.0 AI-027-20260812-global-ai-ne-39241847 Full article

From the mine to the grid

How Komatsu machines help build Atlassian Williams F1 Team’s FW48

Skan AI raises $63M to give AI agents a map of enterprise work
SiliconANGLE AI 2026-08-12 22:57 UTC Score 44.0 USR-0127-20260812-global-ai-ne-09707cb5 Full article

Skan AI raises $63M to give AI agents a map of enterprise work

Process intelligence company Skan AI said today it raised $63 million in a Series C round to help further develop a platform that records how enterprise work actually gets done and feeds that record to artificial intelligence agents. Founded in 2019, the company offers software that sits on employee desktops, grabs screenshots and then processes […] The post Skan AI raises $63M to give AI agents a map of enterprise work appeared first on SiliconANGLE .

The Best Photos of the Big August Solar Eclipse
WIRED AI 2026-08-12 22:47 UTC Score 49.0 AI-015-20260812-global-ai-ne-8785e944 Full article

The Best Photos of the Big August Solar Eclipse

It’s been a century since the Iberian Peninsula has been in the full shadow of the moon. Here’s what it looked like in the path of totality.

Feature Request: Age-Verified Adult Mode for Creative Writing and Long-Term Roleplay
OpenAI Community 2026-08-12 22:43 UTC Score 48.0 AI-116-20260812-social-media-829c9f2b Full article

Feature Request: Age-Verified Adult Mode for Creative Writing and Long-Term Roleplay

I would like to request an optional adult mode for age-verified users, especially for creative writing and long-term roleplay. I fully support strong protections for minors. However, I believe there should be a clearer distinction between underage users and verified adults, so that adults can choose a broader range of mature fictional expression without weakening protections for younger users. My main use case is long-term character roleplay and collaborative fiction. Over time, ChatGPT can build extremely nuanced character dynamics, emotional continuity, relationship history, and consistent characterization. That is one of its greatest strengths. However, when a fictional relationship becomes more physically intimate, the available expression suddenly becomes much more restricted. This creates an unusual user experience: the character writing, emotional nuance, and relationship development may be excellent, but the story must either fade out abruptly or be moved to another service. Moving to another model or platform often means losing the carefully developed characterization and relationship continuity that made the roleplay compelling in the first place. For adult users who have completed age verification, I would appreciate an optional mode that allows a wider range of consensual mature fictional content while still maintaining clear safeguards against harmful, illegal, exploitative, or underage sexual content. A system like this could include: Strict age verification An…

Vibe coding startup Lovable doubles valuation to $13.3B with $400M raise
SiliconANGLE AI 2026-08-12 22:33 UTC Score 43.0 USR-0127-20260812-global-ai-ne-38181d05 Full article

Vibe coding startup Lovable doubles valuation to $13.3B with $400M raise

Lovable Labs Inc. said today it has raised $400 million in Series C funding at a $13.3 billion valuation. The Swedish artificial intelligence coding startup was worth $6.6 billion in December. The company sells a vibe coding service. Users describe what they want in plain language and the platform builds it. Hosting is included. Chief […] The post Vibe coding startup Lovable doubles valuation to $13.3B with $400M raise appeared first on SiliconANGLE .

German cabinet approves new spying rules
Semafor Technology 2026-08-12 22:31 UTC Score 52.0 USR-0094-20260812-global-ai-ne-957d5485 Full article

German cabinet approves new spying rules

The changes augur a “revolution” in Berlin’s intelligence infrastructure, the German interior minister said.

Eli Lilly tries to stop GLP-1 copycats
Semafor Technology 2026-08-12 22:27 UTC Score 52.0 USR-0094-20260812-global-ai-ne-b226efe4 Full article

Eli Lilly tries to stop GLP-1 copycats

Eli Lilly said it filed six lawsuits against US companies it accused of selling illicit versions of its experimental obesity drug retatrutide.

Feature Request: Deferred / Opportunistic Compute
OpenAI Community 2026-08-12 22:25 UTC Score 45.0 AI-116-20260812-social-media-0a7799ac Full article

Feature Request: Deferred / Opportunistic Compute

I’d like to suggest an optional “Process when compute is available” mode for ChatGPT. When sending a message, users could choose between: Process now , normal response priority. Process when compute is available , I’m happy to wait; process the request whenever capacity allows. This would be different from Scheduled Tasks. The user wouldn’t specify when the task should run, they’d simply give OpenAI permission to process it whenever doing so is most efficient. I think this could help OpenAI make better use of available compute, particularly for non-urgent tasks such as research, document analysis, coding, summarization, and other long-running requests. As an incentive, deferred requests could potentially count less against usage limits or receive some other usage benefit. The user is effectively trading response speed for flexibility in when OpenAI uses its compute resources. In short: I don’t need my answer immediately → OpenAI gets more flexibility in scheduling compute → I receive a small usage benefit in return. This could be completely optional, available to both free and paid users, and especially useful for people who are happy to let non-urgent workloads run whenever capacity is available.

Techcrunch 2026-08-12 22:24 UTC Score 33.0 USR-0001-20260812-global-ai-ne-1a9b4529 Full article

AI nuclear power firm Fermi finally has a new CEO

Lee McIntire, an independent member of Fermi's board, has been hired as CEO, more than three months since the company fired co-founder Toby Neugebauer from the top post.

TWIML AI Podcast 2026-08-12 22:18 UTC Score 50.0 AI-148-20260812-podcasts-and-a4d0ac39 Full article

Why Image Generation Needs More Than Bigger Models with Fatih Porikli - #773

Text-to-image models have become remarkably good at producing realistic images. But realism isn’t the same as correctness. Ask for several distinct people, a specific composition, or a high-resolution image generated locally, and today’s models still struggle in surprising ways. In this episode, Fatih Porikli, Vice President of Technology at Qualcomm, joins me to discuss what remains unsolved in image generation and several approaches his team presented at CVPR to address those challenges. We explore why better training objectives can improve controllability, how separating scene planning from rendering may lead to more reliable image generation, techniques for generating 16-megapixel images efficiently on edge devices, and new methods for eliminating the visible artifacts that often appear in AI-powered image editing. Along the way, we discuss reinforcement learning for image generation, agentic image generation pipelines, on-device AI, and what the next phase of progress in generative vision systems is likely to look like. 🗒️ Full show notes: https://twimlai.com/go/773

GitHub Engineering 2026-08-12 22:17 UTC Score 39.0 USR-0062-20260812-ai-specialis-ea7501e3 Full article

GitHub availability report: July 2026

In July, we experienced eight incidents that resulted in degraded performance across GitHub services. The post GitHub availability report: July 2026 appeared first on The GitHub Blog .

InfoWorld AI 2026-08-12 22:14 UTC Score 63.0 USR-0126-20260812-global-ai-ne-e0ac8bc3 Full article

Lovable reaches $13.3B valuation as it adds Cerebras, enterprise tools

Vibe-coding website company Lovable has raised $400 million in Series C funding at a $13.3 billion valuation. The company also recently announced a partnership with AI infrastructure provider Cerebras to accelerate AI inference on its platform. Lovable is a vibe-coding website where users can create full-stack web applications without coding expertise by describing what they want in plain English. The platform combines AI coding tools, real-time collaboration, and project sharing. Customers include the likes of Adidas, Deutsche Telekom, NVIDIA, Udacity, and Workday. In the August 12 funding announcement , the company also unveiled several new Lovable platform capabilities: Built-in payment functionality powered by Paddle and Stripe SEO and AI-search tools to improve discoverability, including integration with Semrush Deeper integrations with Google Workspace, Microsoft 365, Salesforce, Stripe, and ElevenLabs Automatic and scheduled security scanning Additional governance and visibility features including publishing controls, abandoned app clean-up, and workspace insights A dedicated security page, showing which security controls are live for each app In addition, Lovable recently became the first AI coding platform to receive AIUC-1 certification . AIUC-1 is a security, safety, and reliability standard built specifically for AI agents, based on input from Stanford, MIT, MITRE, and the Cloud Security Alliance. Lovable’s $400 million in Series C funding was led by Menlo Ventur…

[8월12일] 저커버그, 폐쇄형 AI 비판했지만…“메타도 오픈소스는 아니다”
Korea AI Times 2026-08-12 22:00 UTC Score 40.0 USR-0048-20260812-global-ai-ne-1458c341 Full article

[8월12일] 저커버그, 폐쇄형 AI 비판했지만…“메타도 오픈소스는 아니다”

마크 저커버그 메타 CEO는 10일(현지시간) 오픈웨이트 모델인 \'뮤즈 글리머\'를 출시하며 장문의 글을 올렸습니다. 6500단어에 이르는 분량으로, AI 경쟁에 대한 회사의 입장과 사회에 미치는 영향, 그리고 미국 정부가 이 기술을 어떻게 규제하고 장려해야 하는지에 대한 내용입니다.핵심은 AI의 가장 큰 위험으로 ‘중앙집중’을 지목하며 오픈웨이트 AI를 적극 옹호했습니다. AI를 소수 기업이나 정부가 독점해서는 안 되며, 개인이 직접 AI를 통제할 수 있어야 한다는 주장입니다.물론, 이는 이전부터 그가 계속 주장하던 바와 크게 다르

Pixel 11 event live blog: Let’s watch Trevor Noah introduce Google’s new phones
The Verge AI 2026-08-12 21:30 UTC Score 49.0 AI-016-20260812-global-ai-ne-76b36feb Full article

Pixel 11 event live blog: Let’s watch Trevor Noah introduce Google’s new phones

It's almost time for the Made by Google keynote, where the company will show off the brand-new Pixel hardware it announced today. Like last year, it'll be a celebrity-packed live show, though Trevor Noah is hosting instead of Jimmy Fallon. If the 2025 show was any indication, today's broadcast might be something of a cringefest. […]