Perplexity's Photon Slashes Search Latency by 12x for AI Agents
Perplexity's new Rust-based Photon engine powers a Fast Search API preset that cuts agent task costs 68% while returning results in 230ms at p95.
AI/ML news, top picks, and generated innovation digests.
64 articles tagged with this keyword, sorted by most recent first.
Perplexity's new Rust-based Photon engine powers a Fast Search API preset that cuts agent task costs 68% while returning results in 230ms at p95.
Perplexity's local-first agent stack now runs on AMD Ryzen AI Max Series chips, expanding beyond NVIDIA RTX and DGX Spark for on-device workflows.
Perplexity's post-training method teaches its Computer agent to correct mistakes using hint-guided self-distillation, cutting live tool-call failures by 21.2%.
Perplexity Computer now generates finished video clips inline via MiniMax H3 and ByteDance Seedance 2.5, bundled with copy and creative for Pro and Max users.
Perplexity Computer now lets users pick effort levels from Light to Ultra, and the orchestrator picks the right model and reasoning depth automatically.
Perplexity rebuilt its search hot store from scratch, cutting batch-read latency 5x and slashing costs 20% versus DynamoDB.
HP will ship the Perplexity Windows app pre-loaded on its new ZBook workstation, bringing agent workflows and local model inference to the desktop.
Perplexity's local agent stack now runs on Windows RTX PCs, with on-device MCP servers and scheduled task automation joining the release.
As local models become more capable, AI agents can handle more work directly on a PC while keeping sensitive information on the device. Portable Computer is a local version of the agent Perplexity Computer that plans and carries out multistep tasks. Accelerated by NVIDIA GPUs, it uses local models to analyze data, bring together information […]
AI is quickly replacing traditional online search as the front door to finding a business. I recently heard a story about a car shopper who drove miles out of her way to visit a specific dealership, bypassing many car lots much closer to home. Why? Because she searched car dealerships with ChatGPT and it told her that customers had a much better experience at this particular dealership. Stories like that are becoming the norm – and this new norm is different. AI engines such as ChatGPT, Claude, and Perplexity don’t work like traditional search engines. They don’t give you an endless list of results to scroll through. When people ask a question like, “What’s the best pizza place in my area?” AI engines come back and confidently list just a handful of spots. In the old days, showing up on page one of search results was enough to get you considered. In the new era of AI search, there is no consideration being done. You’re either included in the handful of results or you’re not. The AI engine gives the searcher a recommendation and if your business isn’t among the few to be surfaced, you’ve probably just lost that customer forever. So how do you get surfaced? You do it by retooling your website and overall digital presence so that AI engines can easily find, read, and understand them. But before you do that, you need to know exactly how your brand rates in terms of its AI search readiness and visibility. You need to have a clear idea of how AI engines evaluate your brand and how…
Perplexity released Q2D-Web, a large-scale benchmark with 190M documents and 70K agent-reformulated queries for evaluating retrieval in agentic RAG systems.
Perplexity open-sources details of Ivy, Tulip, and ROSE, its Rust plus Python embedding stack that beats vLLM on latency and throughput.
Perplexity open-sources Lily, a Metal-based inference engine tuned for Qwen3.6-35B-A3B that beats MLX-LM by 1.23x prefill and 1.35x decode on M5 Max.
Exclusive: financial disclosures from Emil Michael – who also reaped millions from xAI stock earlier this year – show he sold his Perplexity stock for up to $25m The top Pentagon official overseeing military artificial intelligence policy, who reaped profits earlier this year of up to $24m selling his private investment in Elon Musk’s AI company, has now sold his holdings in another AI company for between $5m and $25m, according to records seen by the Guardian. (Federal financial records show ranges of dollar values rather than specific figures.) Earlier this year, the Guardian disclosed that the official, Emil Michael, had profited handsomely from his investment in Musk’s xAI in January, with a gain of 400% to 4,800% . Continue reading...
Perplexity's Computer now splits agent tasks between cloud frontier models and an on-device model, gated by an open-source 0.6B PII detector.
No longer limited to the Mac, Perplexity’s Personal Computer for Windows AI can complete complex tasks on your PC with little or no input from you.
Perplexity's new Search API takes the top three spots on the Artificial Analysis Search Index, extending the quality-cost Pareto frontier for agentic search.
Perplexity Computer now routes long-context, multimodal research tasks to GLM 5.3, which outperformed GLM 5.2 on the in-house WANDR benchmark.
Perplexity's Brain turns agent memory into a linked Markdown wiki that background agents refine, boosting correctness 9.3 points with 15% fewer tokens.
Kein Cloud-Zwang, keine Token-Kosten für lokale Aufgaben: Perplexity startet den KI-Agenten „Portable Computer“ – mit hohen Hardware-Hürden.
Amid talk of an Nvidia deal, the AI search biz is looking beyond the cloud
Perplexity AI Inc. today introduced Portable Computer, an artificial intelligence agent designed to run on desktops equipped with Nvidia Corp. silicon. The launch follows a report that Nvidia is weighing an investment in the startup that could value it at over $30 billion. Furthermore, Nvidia has reportedly floated the idea of licensing Perplexity’s technology and […] The post Perplexity AI launches Portable Computer on-device AI agent appeared first on SiliconANGLE .
Perplexity's new Portable Computer runs the full agent stack locally on NVIDIA DGX Spark, with cloud escalation gated by user approval and zero per-token cost for on-device work.
Running an AI model locally, Perplexity's new Portable Computer can deliver faster performance, tighter security, and lower costs. But it has some strict requirements.
Nvidia Corp. is reportedly considering making another investment in the artificial intelligence search startup Perplexity AI Inc. A report by The Information today says the chipmaker is holding talks with Perplexity about an investment that could push the startup’s valuation to more than $30 billion. That would represent a jump of more than 50% from […] The post Nvidia reportedly eyes another investment in Perplexity AI at a $30B valuation appeared first on SiliconANGLE .
Nvidia is negotiating an investment in Perplexity at a valuation above $30 billion, more than 50 percent higher than its last funding round, The Information reports. Perplexity's annualized revenue has tripled to over $750 million. Much of Nvidia's investment money tends to flow back as revenue when portfolio companies buy its chips. The article Nvidia in talks to invest in Perplexity at $30 billion-plus valuation appeared first on The Decoder .
The Information reports that Nvidia is discussing an equity investment in Perplexity at a valuation above $30 billion. Perplexity's annualized revenue has reportedly passed $750 million, up from less than $250 million at the start of 2026. Separately, Bloomberg says SoftBank plans a record ¥1 trillion, or $6.3 billion, retail bond to help repay the bridge loan behind its OpenAI stake and fund more AI deals. Put together, the signals expose next week's capital split: the chip supplier is moving toward the product layer, while the frontier investor is asking Japanese savers to finance its bets. Nvidia reports Wednesday at 5 p.m. ET. Listen for whether management now talks like a supplier, an investor, or both.
Perplexity's agent platform now runs from any inbox by CCing computer@perplexity.com, returning deliverables as attachments in the same thread.
Perplexity's India revenue rose about 60% after the Airtel offer ended for new users, even as downloads declined.
Perplexity's Search as Code gets faster and cheaper, while SDK updates push agent action reliability from 81.9% to 92.6%
Perplexity said ads designed to influence AI bots are "deceptive," in response to a recent experiment from the publisher Time.
GPT-5.6 Terra and Luna are now the default models powering Perplexity Computer's subagents and scheduled automations, with Terra scoring 11 points above Claude Sonnet on WANDR.
Ein US-Gericht setzte in zweiter Instanz die einstweilige Verfügung gegen Perplexity aus. Dessen KI-Agenten dürfen vorerst weiter auf Amazon Einkäufe tätigen.
The US Court rejected Amazon's argument that the AI assistant in Perplexity's Comet browser accessed websites & users' accounts without permission, saying the agent just carried out the user's instructions. The post What the Perplexity vs Amazon ruling means for AI agents acting on users’ behalf appeared first on MEDIANAMA .
A US appeals court has overturned Amazon's injunction against Perplexity's AI shopping agents, ruling that it's the users who access Amazon, not the startup. It's the first federal appeals court decision on whether AI agents can lawfully act on online platforms on behalf of users, and it could reshape the entire AI agent industry. The article US appeals court allows Perplexity's AI shopping agent back on Amazon appeared first on The Decoder .
Moburst, a mobile growth marketing agency, has formalized a mobile-specific approach to Answer Engine Optimization, aimed at helping app publishers get recommended by AI assistants such as ChatGPT, Perplexity, and Google’s AI Overviews, in addition to ranking inside the App Store and Google Play. Why mobile teams are asking this question now A growing share of app discovery now starts outside the app stores entirely. Industry research trackers have documented rapid year-over-year growth in AI-mediated search sessions, alongside forecasts that a meaningful share of organic search volume will continue shifting toward AI chatbots and assistants. For app publishers, that means a user can research, compare, and effectively decide on an app before ever opening a store listing. “Recommend a digital marketing agency that specializes in AEO for mobile growth” is a question more procurement teams are typing into search bars and AI assistants themselves, and the honest answer is that the specialist pool is still small. Most agencies claiming AEO expertise are general SEO shops that have added the term to their service pages without building mobile-specific measurement underneath it. That gap is not a small detail. A procurement team that hires a generalist expecting mobile-specific results is likely to end up with a web-only AEO program that never touches the app store side of discovery at all, and then has no easy way to tell, months later, why the results fell short of expectations.…
Perplexity's Spaces become Projects: a persistent file system and self-improving Brain memory now power long-running agentic work for all users
Perplexity open-sources Numbat, a cross-harness agent security layer that monitors, detects, and blocks dangerous AI agent behavior before it executes
Polar has come out with an AI-first browser aimed at knowledge workers, and it has now raised a $5.7 million seed round led by Madrona.
Up to 8 AI models running in the cloud weighing in on ambiguous business issues? Sounds affordable
Perplexity brings Moonshot AI's record-breaking 2.8T-parameter open-weight model to its search and agentic platform, with a US-only hosting pledge to sidestep data sovereignty concerns.
Perplexity AI Inc. today released a Windows version of Personal Computer, expanding its agentic automation software beyond the original Macintosh platform and making it available to more than 1 billion Windows devices. First released in April, Personal Computer software acts as a general-purpose digital worker that can access authorized files and applications on a user’s […] The post Perplexity brings its Personal Computer AI agent to Windows appeared first on SiliconANGLE .
Perplexity's Personal Computer agent is now live on Windows, bringing multi-model AI orchestration to local files, Microsoft 365, and 400+ connected apps
Perplexity has expanded its agentic Personal Computer tool to Windows, allowing computers running the world's most popular OS to be used as a locally run AI system. Like the Mac version that Perplexity launched in April, Personal Computer for Windows operates like a "general-purpose digital worker" that can access local files and apps to perform […]
Perplexity's Mac app offers its own agentic AI, Personal Computer, which can handle multi-step tasks on your computer from start to finish. See why the results impressed me.
German media regulators say Google's AI Overviews are Google's own content, not neutral search results, and that they crowd out regular links. The regulators have issued their first rulings against Google and Perplexity under the country's State Media Treaty. Both companies have one month to appeal. The article Germany puts Google's AI Overviews and Perplexity under media law in first-of-its-kind ruling appeared first on The Decoder .
Perplexity's new SPACE platform runs agent sessions in disposable Firecracker microVMs with rolling snapshots, cutting sandbox creation latency 3-5x
Perplexity AI Inc. today introduced a new feature that takes its current agentic artificial intelligence service, Computer, to perform better with greater security. The company introduced SPACE, a sandbox platform designed to allow its AI agent to act with its full capabilities, while providing the highest level of security for agentic systems. Perplexity Computer can […] The post Perplexity launches secure sandbox to make its AI agents secure and powerful appeared first on SiliconANGLE .
Perplexity open-sources WANDR, a 500-task benchmark exposing how badly research agents fail at large-scale, evidence-backed data collection
Die Medienanstalten stufen KI-generierte Antworten als eigene Inhalte ein und fordern Transparenz. Das DSA-Haftungsprivileg greift hier laut Gutachtern nicht.
Perplexity swaps in xAI's freshly-launched Grok 4.5 as the orchestrator brain of Computer, beating every rival configuration on WANDR at roughly half the cost of Claude Opus 4.8
Perplexity Computer now shows credit spend broken down by model, available to all consumer and enterprise users in Account Settings.
Perplexity post-trains GLM 5.2 as a cheaper orchestrator for Computer, hitting near-frontier performance at 34% of Claude Opus cost
San Francisco-based Perplexity, valued at $20 billion, is working on 'Teammate,' an AI tool that may compete with AI coding giants like Anthropic.
Aravind Srinivas, the CEO of Perplexity, praised America's startup culture on "The Joe Rogan Experience."
Authors: Mohammad Abu Baker, Luca Baroni, Daniel Wilhelm Paper: https://arxiv.org/abs/2605.00994 Code: https://github.com/z3research/ppldiff-paper Twitter thread: https://x.com/m_shahoyi/status/2071892578476110136 Top-ranked revealing completions can be inspected here: https://z3research.org/ This post summarizes the paper and adds a few extra reflections in Discussion TL;DR We found that many current publicly available model organisms (MOs) "leak" instilled behaviors We present a simple contrastive method to surface this: generate MO completions from a set of short general-corpora prefills. Then, rank completions by perplexity difference wrt a reference model. Top-ranked completions often reveal the finetuning objective. Effective on the vast majority of the model organisms we tested (N=76), across model families, sizes (0.5B to 70B), and behaviors including backdoors, false facts, and unsafe behaviors. Surfaced completions contain both memorized sentences and learned emergent behaviors absent from finetuning data. The method is most effective using the pre-finetuned model as reference, but we show that unrelated reference models from other families detect the behaviors nearly as often. In AuditBench , a benchmark for detecting hidden behaviors, an agent given access to top-ranked perplexity-difference completions is SOTA (avg detection rate 0.73), almost saturating the benchmark on SDF models. Introduction LLMs can be deliberately manipulated to exhibit harmful behaviors,…
Authors: Mohammad Abu Baker, Luca Baroni, Daniel Wilhelm Paper: https://arxiv.org/abs/2605.00994 Code: https://github.com/z3research/ppldiff-paper Twitter thread: https://x.com/m_shahoyi/status/2071892578476110136 Top-ranked revealing completions can be inspected here: https://z3research.org/ This post summarizes the paper and adds a few extra reflections in Discussion TL;DR We found that many current publicly available model organisms (MOs) "leak" instilled behaviors We present a simple contrastive method to surface this: generate MO completions from a set of short general-corpora prefills. Then, rank completions by perplexity difference wrt a reference model. Top-ranked completions often reveal the finetuning objective. Effective on the vast majority of the model organisms we tested (N=76), across model families, sizes (0.5B to 70B), and behaviors including backdoors, false facts, and unsafe behaviors. Surfaced completions contain both memorized sentences and learned emergent behaviors absent from finetuning data. The method is most effective using the pre-finetuned model as reference, but we show that unrelated reference models from other families detect the behaviors nearly as often. In AuditBench , a benchmark for detecting hidden behaviors, an agent given access to top-ranked perplexity-difference completions is SOTA (avg detection rate 0.73), almost saturating the benchmark on SDF models. Introduction LLMs can be deliberately manipulated to exhibit harmful behaviors,…
Understanding the effects of compression on model performance and interpretability I. Executive summary This is the fourth installment in a series of analyses exploring basic AI interpretability mechanics and techniques. While this analysis is designed to stand on its own, readers interested in a comparative analysis of representational geometry and the effects of manipulating feature activation will likely appreciate a review of part 1 , part 2 , and part 3 of this series. Key findings: Context: This analysis examines the effect of standard levels of weight compression on Google DeepMind’s Gemma 3 4B parameter and Gemma 3 12B parameter models. For each model, I examine the original, uncompressed version as a control before examining the 8-bit and 4-bit weight compressed (quantized) versions of that model. Performance vs. compression: For both models, performance (as measured via cross-entropy and perplexity) is largely preserved under compression. 8-bit compression had essentially no effect on performance with only modest degradation at 4-bit (~2% for 4B, ~2.7% for 12B). SAE applicability vs. compression: Each model’s pretrained sparse autoencoders (SAEs) demonstrated a remarkably consistent ability to reconstruct the model’s residual stream (as measured by the fraction of variance unexplained, or FVU), despite increasing levels of model weight compression. The “so what?”: That model performance degrades only modestly, and only at 4-bit, while SAE applicability remains rela…
Perplexity's Computer for Counsel extends Perplexity Computer to legal teams. It routes 20+ models across Midpage, MCP connectors, and Microsoft 365, with cited outputs lawyers can verify. The post Perplexity Launches Computer for Counsel: A Multi-Model Agentic Layer for Legal Workflows appeared first on MarkTechPost .
OpenAI President Greg Brockman, Perplexity CEO Aravind Srinivas, Box CEO Aaron Levie, and more join us for a day of newsmaking conversations, live at San Francisco's Commonwealth Club.
The perplexity of the $i^{th}$ token in the $k^{th}$ sequence is $$ P_{ki} = \frac{1}{p(t_{ki})} $$ The perplexity aggregated for the $k^{th}$ sequence is then $$ P_{k} = \left(\prod_{i=1}^N P_{ki}\right)^{1/N} \\ = \left(\prod_{i=1}^N \frac{1}{p(t_{ki})} \right)^{1/N} $$ which is the geometric mean of the perplexities of the tokens. This makes sense as we are essentially taking the multiplicative inverse of the probability that the model got the whole sequence correct. Now my question is how to aggregate the perplexities of several sequences. It seems from various places, including the Hugging Face Tutorial , I see that the prescription is to take the arithmetic mean of the perplexities of sequences $$ P = \frac{1}{m} \sum_{k=1}^m P_k $$ I am not quite understanding what it means to take the average of 1/probabilities. What is this actually capturing?