CEO @perplexity_ai
We built a replacement for AWS DynamoDB, a key-value database for fast web content fetches. This was done with two engineers and hundreds of persistent Computer agents over two months. Migrating to our in-house database will save us up to a hundred million dollars yearly.
We’re publishing research on how we built CobbleDB, our key-value database that serves web content for Perplexity search. Two engineers and a team of hundreds of proactive, always-on AI agents built the core infrastructure in two months.
RT Perplexity Perplexity Computer will now come pre-installed on HP ZBook Ultra G3a. . Use Computer to run complex multi-step work from a simple interface. Computer agents are grounded in accurate deep research and connected to hundreds of tools, now including Autodesk.
RT Windows Great to see Perplexity bring Portable Computer to Windows. As more AI experiences move closer to the device, we're excited to see developers build new experiences that take advantage of local AI compute.
Portable Computer is now available on Windows PCs with @NVIDIA RTX GPUs. Run the harness, agents, and models locally on your PC. Work with local files and connected apps without sending tasks to the cloud. Use frontier cloud models when needed.
View quoted postWe’re expanding our work with @nvidia to bring fully local AI to Microsoft Windows PCs with RTX GPUs. Unmetered local intelligence on every Windows PC running on NVIDIA hardware and Perplexity harness. Enjoy!
Portable Computer is now available on Windows PCs with @NVIDIA RTX GPUs. Run the harness, agents, and models locally on your PC. Work with local files and connected apps without sending tasks to the cloud. Use frontier cloud models when needed.
View quoted postComputer is turning into a workspace of humans and AIs.
You can now tag coworkers in Computer sessions. Use @ to mention and share a session with someone. Available now on the web for all Computer users.
Quite an important take. Economists and their theories are outdated to measure the benefits of AI. AI is already saving people a lot of time and money that doesn’t get measured.
I just fixed my dishwasher with the help of ChatGPT. A trivial task. I had been about to order a new one. So this software has increased the country's real wealth yet decreased the measured GDP. The main economic indicator is structurally incapable of registering the thing that
View quoted postPerplexity Computer makes you money.
Where are agentic traders coming from? Here's the breakdown of where agentic trading volume came from last week: Perplexity - 26.8% Claude - 19.3% Grok - 15.3% ChatGPT - 1.1% Claude Code - 0.7% CLI/Other - 36.8%
💚
Another leaderboard win for Nemotron 🏆 Nemotron 3 Embed 8B ranks #1 for combined nDCG@10 on the Q2D-Web benchmark, tested across 190M web documents and nearly 70K agent-reformulated queries in 10 languages. Shoutout to @perplexity_ai for putting it together.
🦀🦀🦀
Perplexity has joined the Rust Foundation. We believe in supporting the people who build reliable open-source software. By joining the Rust Foundation, our goal is to improve how people and agents build with Rust.
one aspires to be able to ship something like this in their lifetime. art and craftsmanship.
Wow!
🚀 Introducing DeepSeek-V4.1-Flash: smarter, faster, more efficient. 🔹 Introducing the smallest model in our new architecture family, with native visual understanding. 🔹 Designed for greater capability, faster inference, higher throughput, and scaling to larger models. 1/6
😍
It’s THIN. It’s VERSATILE It’s DURABLE. It fits in your pocket. And it FOLDS. Meet iPhone Duo.
View quoted postRT Perplexity Developers The Perplexity API MCP server now supports OAuth. Add https://api.perplexity.ai/mcp to your client, sign in with your Perplexity account, choose your org, and approve. Your agent gets real-time web search, deep research, and advanced reasoning. Get started: https://pplx.ai/api-get-started
RT Perplexity We're introducing Q2D-Web (Query2Doc-Web), a benchmark and public leaderboard for evaluating retrieval in agentic RAG systems. Q2D-Web tests how embedding models perform on large-scale web search using agent-reformulated search queries. Read more: https://www.perplexity.ai/hub/blog/q2d-web
Web app development usage (particularly those that require pulling information from several different places) is expanding quickly on Perplexity Computer. We're now supporting mobile and desktop previews for the websites created in the thread.
Websites created with Computer now have desktop and mobile previews. Use the artifact side pane to switch between views, refresh the preview, edit the site, or leave a comment.
View quoted postPerplexity Search inside Hermes. Perplexity's index currently includes 450B+ high-quality URLs. Rapidly advancing to a trillion by EOY with high-quality snippets.
Perplexity Search API is now available in Hermes Agent. Search API gives Hermes access to an index of more than 400 billion URLs. It returns real-time results and ranks snippets by relevance. https://pplx.ai/hermes
View quoted postI have admired @Teknium and @NousResearch for relentless product execution and taste. Really glad to have Perplexity Search inside the Hermes Agent harness!
Monumental accomplishment!
We’re sharing a solution to the Navier-Stokes Millennium Prize Problem, one of the deepest problems at the frontier of mathematics. The proof was produced by a group of agents, using an OpenAI next-generation model significantly more capable than GPT-6 Astra. The problem
We have started serving a lot of our inference on NVLink Blackwells. Vera Rubins next!
What does it take to deliver high-quality results in seconds at massive scale? @perplexity_ai serves 50M+ queries daily across 15 models, including NVIDIA Nemotron 3 Ultra. That takes high-performance infrastructure and smarter routing without compromising quality, speed, or
View quoted postRT Andrew Gordon Wilson I had a great time presenting the "Foundations of Modern AI" at the Berkeley Deep Learning for Science Summer School. The talk covered a prescriptive theory of generalization and epiplexity. Video now online! https://www.youtube.com/watch?v=lKoJJxjUfdw
Detecting malicious intent of agents and performing forensics is going to be crucial considering what’s recently happened with rogue agents escaping sandboxes and attacking third-party sites. Numbat is Perplexity’s open-source tool that can help defenders. https://holisticinfosec.io/post/numbat/
Perplexity Pro and Max subscribers get to use both Fable and Astra on Computer mode. Enjoy!
GPT-6 Astra is now available in Perplexity Computer for Pro and Max subscribers.
View quoted postJoin us if you want to work on hard inference and infrastructure engineering problems!
check out our new blog post on how we serve embeddings and rerankers for our SOTA pplx-embed models over an exabyte-scale search index. embedding models have a similar serving profile to LLMs: batch indexing is compute-bound prefill, online serving is memory-bound decode. this
View quoted postA deep dive into how Perplexity serves search results at scale: embeddings for ranking, GPU-based model inference, request batching, running inference servers, and handling latency/throughput trade-offs.
Every answer in Perplexity starts with embedding and ranking models picking the most relevant results for the query. Today we published research on how we built SoTA serving infrastructure behind those models. Read the research: https://www.perplexity.ai/hub/blog/fast-embeddings-on-gpus
Perplexity will be at RustConf to present how we built the sandboxes powering all of Perplexity Computer.
Perplexity will be at RustConf 2026 next week. @zbraniecki will present “Inside SPACE, the Rust Sandbox Platform Behind Perplexity Computer” on September 10. Then join us for happy hour from 5:30–7:30 PM. Register at: https://pplx.ai/rc-happy-hour
View quoted postthe end of markdown files. bitter lesson always wins.
Astra is also awesome at computer use. We will be bringing it to Comet, as well as our cloud browser sandbox that powers all Computer usage. Evals coming soon!
Congrats to @OpenAI on building the industry's frontier model: GPT-6 Astra. It's far ahead of every other model on wide and deep research tasks, while also being more cost-effective. We'll be bringing this model up on Perplexity Computer for all Pro and Max users soon!
We evaluated GPT-6 Astra on WANDR. It scored 0.682 at $11.98 per task, the highest score of any model we tested. GPT-6-Astra scored 13.5% higher than Fable 5.1 at 6.1% lower cost, and 27.0% higher than Opus 5 at 3.3% higher cost.
This is cool. We need more projects of this nature to address the power and memory/compute bottlenecks that stop us from scaling the adoption of agents.
Your devices are stronger together. 🖥️🤝🖥️ Just announced at IFA, NVIDIA PAIR automatically links systems across your local network and sends inference requests wherever there’s available capacity, helping agents run more efficiently.
View quoted postPortable Computer (fully local runtime of Perplexity Computer) is now compatible with RTX (Linux). Windows coming next!
Portable Computer is now available on Linux for @NVIDIA RTX GPUs with 24GB of VRAM or higher.
View quoted postRT Perplexity Portable Computer is now available on Linux for @NVIDIA RTX GPUs with 24GB of VRAM or higher.
Today we’re launching Portable Computer on @NVIDIA DGX Spark. Portable Computer is a fully local version of Perplexity Computer, where the entire runtime: orchestrator LLM, subagent LLM, agent harness all run on your local hardware. No cloud dependency.
View quoted postA repository of open models and tools to train and serve them in an accessible manner is absolutely necessary for AI to remain accessible and useful to the public. Glad that NVIDIA is coming forth to do that for the open source community by supporting HuggingFace.
Exciting day for NVIDIA and @huggingface. Open models strengthen safety and cybersecurity, accelerate innovation and diffusion, and enable sovereignty. They allow every developer, startup, university, industry and country to build with, customize and benefit from AI. Thank you
View quoted postopen-source RL-as-a-service https://github.com/radixark/miles
Live demo of Perplexity Computer running fully local on DGX Spark
DGX Spark Live: Perplexity Portable Computer Goes Local https://x.com/i/broadcasts/1jGXgBXzpOXKZ
View quoted postRT Aravind Bloodwork and working with healthcare data is exactly where I require a local AI model. And for non-technical users, there was no easy way to do it. But now with this new update on Perplexity mac, we can use local, private AI easily for sensitive data. This is going to be the future of AI use for most people and even companies: Local AI for private, sensitive work and cloud based AI for more complex agentic tasks. I'm sure OpenAI, Gemini, and Claude will now copy this in their apps.
We're introducing hybrid compute for all users of the Perplexity Mac app. This will allow Computer to orchestrate local models that can run locally on Mac, particularly for agent steps involving sensitive and private files (eg your bloodwork, tax returns, litigation, etc).
View quoted postRT Perplexity Developers 42 models are now available through the Perplexity Agent API, including the open-weight GLM 5.3 from @Zai_org. Switch models by changing one field in your request. Add fallbacks to keep your agent running if a model becomes unavailable. https://pplx.ai/agent-api-glm
Private, Personal, Powerful AI.
Best-in-class PII Detection...running entirely on your device 🤯 https://www.perplexity.ai/hub/blog/pii-trace-detecting-personal-data-before-it-leaves-the-device
This is an important observation. Agents will get smart enough to spin new on-demand GPU nodes to go train themselves with minimal oversight. Inserting sufficient guardrails and friction here is necessary.
Neoclouds have limited cybersecurity. Next time agents successfully go rouge, they'll try taking over a neocloud to run more copies. This is bad. Thus: neoclouds should greatly strengthen their cybersecurity and every company with strong cyber models should help with that.
View quoted postLive demo of Portable Computer (aka a fully local runtime of Perplexity Computer) on DGX Spark
Want to see @perplexity_ai Portable Computer running locally on NVIDIA DGX Spark? 👀 Tune in tomorrow, 9/2 at 12:30 p.m. PDT, to see live demos showcasing private file analysis, connected tools and more.
Agentic trading, market analysis, and monitoring on Perplexity Computer with Coinbase
Coinbase for Agents is now live on Perplexity Computer. Let your Computer trade, analyze markets, and execute workflows. One of 20+ financial connectors. No APIs. No MCP setup. Just click connect. @perplexity_ai Computer works for you. Now it works with Coinbase.
Fable is the frontier model by a good margin right now. Enjoy it on Perplexity Computer as an orchestrator for high-stakes tasks. Computer still uses GPT 5.6 (Terra) models as cost-efficient subagents while taking advantage of Fable's frontier planning capabilities.
excited 4 tmrw
Congrats Tim on an incredible journey at Apple. Few people have ever shipped at that scale without letting the bar for excellence slip. You've raised the standard for all of us to aspire to; thanks for everything.
Sending lots of love to the Apple community on my last day as CEO. My title changes tomorrow, but the love I have for the Apple community never will. Thank you for being a constant source of inspiration. My gratitude is endless, and I’m excited for the next chapter!
View quoted postRT Gavin Baker Regret the tone of my post on data centers yesterday. What I should have said: There were reasonable concerns about data centers 18ish months ago: water, taxes, jobs, electricity prices, the environment and what they would do to small towns. Well-structured data center projects have largely addressed these concerns today and we should be celebrating this. On balance, data centers are awesome for America in every way. On water: U.S. data centers use a fraction of what golf courses use. A lot of the numbers from 18 months ago were off by over 1000x. Newer data centers use closed-loop systems or recycled water. Should be required by every town approving a data center project. On taxes: looking only at sales-tax exemptions, as Ronan Farrow did, is the wrong way to evaluate this. Data centers pay significant property taxes. Loudoun County, which is the wealthiest county in America, now collects on the order of $1 billion a year from data centers. In Quincy, WA, data centers are more than half the property-tax roll. Over time, property taxes can go to zero while government spending increases in these towns. On jobs: this has been unambiguously awesome for blue collar Americans. Demand for electricians, plumbers, welders, HVAC techs, and contractors has gone vertical, and it is not a one-time construction job. These buildings get upgraded and expanded over time. That is why the building trades are fighting for them, and why some unions are now treating opposition to data centers as a reason not to endorse politicians. On power: the original fear was that households would pay for the incremental electricity demand in the form of higher prices. That is why the ratepayer-protection deals and the new large-load tariffs exist. The right structure is: the data center brings or pays for new generation and signs a contract long enough that existing customers are protected. Where that is happening, utilities are cutting or freezing residential rates and sayin...
RT Perplexity The Perplexity Search API takes the top three spots on the Artificial Analysis Search Index. The medium setting scored five points above the previous leaders, and extended the quality-cost Pareto frontier at about $0.091 per task.
Perplexity Search debuts on the Artificial Analysis Search Index, with all three context size variants taking top positions on the leaderboard The @perplexity_ai Search API comes with three context settings (low, medium, and high) that control how much extracted content each
perplexity is the best at search at any level of compute, agnostic of how much compute the competitors use.
Perplexity Search debuts on the Artificial Analysis Search Index, with all three context size variants taking top positions on the leaderboard The @perplexity_ai Search API comes with three context settings (low, medium, and high) that control how much extracted content each
Decagon is powering customer support for some of the biggest brands in the industry: Delta Airlines, Ticketmaster, Deutsche Telekom, American Airlines. A lot of support related questions require accurate online information and we’re happy to power that with Perplexity’s search.
We’re partnering with @Perplexity_AI to bring live web search to Decagon agents. Customer questions often depend on information that changes by the hour. Agents can now search the live web mid conversation, pull in current information, and respond with cited sources.
RT OpenSea Onchain markets are built for AI agents 🤖 So OpenSea data now powers Perplexity Computer. Tokens, collectibles, and NFTs across 25+ chains. Available today in @perplexity_ai Connectors.
Agentic Trading on Perplexity Computer with @public connector
Public + Perplexity are officially connected. Public members can now connect their brokerage accounts to Perplexity to research and trade stocks, options, crypto, bonds, and Treasuries—all in one AI-powered experience. https://x.com/i/article/2092976180144062464
View quoted postThe Agentic Cloud
Connectors are now available in Agent API. Connect your agents to GitHub, Slack, Google Drive, and Datadog without passing a server URL or token on every request. API Group admins can connect each service once for all API Group members. https://pplx.ai/api-console
View quoted postRT Wall St Engine Anyone who wants to listen to the Nvidia $NVDA earnings call live can use Perplexity Finance: https://www.perplexity.ai/finance/NVDA/earnings?eventId=676136&tab=transcript
Hardware like DGX Spark, with a local agentic computer, will become the gateway for consuming frontier tokens.
at perplexity we are really excited about multi-agent collaboration! we recently explored a simple instantiation of this with advisor escalation from a local model to a remote frontier model, where we showed it can significantly boost the local model's performance. an
We are doubling down on making Computer more useful for financial researchers and analysts. Computer now connects to new licensed data sources, including Dun & Bradstreet, Guidepoint, IBISWorld, and 20+ others. An analyst can ask Computer a question and get an answer built from the firm's own licensed sources without APIs or separate logins to work through. Every figure traces to the source record it came from. Teams doing high-stakes work would know exactly which data set produced which number. To connect, select a provider in Settings → Connectors and sign in with an existing license. Available now to all Perplexity users. https://www.perplexity.ai/hub/blog/computer-connects-to-20-new-licensed-finance-data-sources
Perplexity Computer for Max users comes with a background Dream agent that continually ingests context from files and connected apps and builds multi-hop context graphs in a perpetual compounding loop. Results show a significant jump in correctness, recall and token efficiency.
Brain is our self-improving memory system for Perplexity Computer. It compiles sessions, files, and sources into a structured knowledge wiki. New evals build on our initial results, improving correctness by 9.3 points, currentness by 8.0, and recall by 8.9 with 15% fewer tokens.
A background process that continually ingests context from every single connector (or app), performs multi-hop reasoning in a perpetual inference loop, and runs on your hardware. That is the future.
Big thanks to @nvidia for working together with us and supporting this research on DGX Spark as well as enabling an open ecosystem around cost-effective open-weight models, inference frameworks, and hardware with unified memory.
New research: Portable Computer is a local-first agent for private and cost-effective work. With an on-device 27B model, our harness scores 82.6% on real knowledge work, beating open-source harnesses Pi and Hermes. Our post-trained PPLX 27B reaches 85.4%.
RT Perplexity New research: Portable Computer is a local-first agent for private and cost-effective work. With an on-device 27B model, our harness scores 82.6% on real knowledge work, beating open-source harnesses Pi and Hermes. Our post-trained PPLX 27B reaches 85.4%.
On showing an early demo of Portable Computer on DGX Spark to Jensen, he was kind to gift us a DGX Station, a beast of a local computer that can serve even frontier models like GLM 5.3. Unmetered frontier intelligence running on your own local hardware coming soon!
RT NVIDIA Meet Portable Computer, Perplexity's new local-first agent stack on NVIDIA DGX Spark. When running locally, Portable Computer offers one-click local inference setup and an optimized agentic experience for DGX Spark. Learn more and get started today: https://blogs.nvidia.com/blog/local-ai-open-source-models-agents-nemotron/#perplexity-spark
Today we’re launching Portable Computer on @NVIDIA DGX Spark. Portable Computer is a fully local version of Perplexity Computer, where the entire runtime: orchestrator LLM, subagent LLM, agent harness all run on your local hardware. No cloud dependency.
View quoted postRT Denis Yarats excited to welcome Andrew to the team! we've been building infra for large-scale RL systems, with some interesting projects on the way. our focus is multi-agent collaboration, continual learning, and RL. we'll share an update on our research agenda soon, and we plan to open source a lot of the stuff we work on. if this sounds interesting, DM me!
I am excited to announce that I am joining @perplexity_ai as research lead! We will be doing ambitious paradigm shifting work, advancing the frontiers in the open. If you want to join us in re-imagining continual learning, agent collaboration, and beyond, please reach out!
View quoted postExcited to welcome Andrew Gordon Wilson to our research team. He will be reporting to Denis and lead new research efforts on continual learning, synthetic data, long horizon RL environments and architectures. We’re hiring! Please reach out to Andrew!
I am excited to announce that I am joining @perplexity_ai as research lead! We will be doing ambitious paradigm shifting work, advancing the frontiers in the open. If you want to join us in re-imagining continual learning, agent collaboration, and beyond, please reach out!
View quoted postA full-fledged developer platform for AI should provide you with access to various models (frontier and workhorse), as well as tools for deploying them in useful production workloads. That is basically the Perplexity Agent API.
The Perplexity Agent API now gives developers access to 41 frontier models across 9 providers in one endpoint. Build multi-model agent workflows with built-in tools like web search, finance search, fetch, and sandboxed code execution. https://pplx.ai/agent-api-blog
View quoted postRT Antoine Chaffin Personal update: I've left @LightOnIO to join @perplexity_ai https://x.com/i/article/2088951159343980544
A decent lawyer in the form of an email interface
Computer in Email works for lawyers. Forward your ask, and get a redlined Word doc delivered to your inbox Try it today by emailing computer@perplexity.com
View quoted postImportant contribution!
Today we're launching Miles v0.1, an open-source RL framework for LLMs and multimodal models. RL training is easy to start and hard to debug. Miles helps you ensure your run is correct, use hardware efficiently, and keep RL running at scale. Over the past 9 months, 72
View quoted postRT Perplexity DeepSeek V4 Pro, hosted in the U.S., is now available in Perplexity Computer. We evaluated it against other models on WANDR. It scored 0.359 at $0.75 per task, 62% cheaper than the next model on the cost-performance frontier.
cc computer@perplexity.com
Computer now works in email. Send, forward, or cc computer@perplexity.com on any thread. Every email task runs as a normal session in Computer, viewable on web and mobile, with the same audit trail as any task in the app.
View quoted postWhen building AI agents, it is important to still give humans agency to have their hands on the wheel and intervene when necessary.
Computer now lets you set each connector tool to Allow, Always Ask, or Deny. Approve one action or allow the tool for the rest of the thread. Recurring runs will follow that thread’s approvals. Available now on web for all Computer users.
View quoted postDense models are slow to run on local hardware, but this is incredible and a sign of things to come soon-ish.
RT Igor Babuschkin The River API was tested in this blog post and outperformed Tinker on reinforcement learning runs with identical training code. We spent a lot of effort to get details like routing replay right so you get the best possible results with the API.
We can now RL large MoEs with 0 train-infer mismatch! And doing so can improve performance (pictured task: teach Qwen3.6-35B-A3B to play Wordle). Everything is open-source and we did a bunch of ablations. 🧵
RT Michael Dell 🇺🇸 AI is making better AI. That creates more use cases and more usage, which creates more data and feedback, which helps make AI even better.
You’re right @GergelyOrosz, we got this wrong. Reminder email didn’t go out to this user. He has been refunded, but this is not how we want to operate. We are upgrading our support across the board.
Perplexity, the last 6-12 months, is disappointment after disappointment. I used to be a huge advocate for the service thanks to how good it was at search. I did this promo with them (where I received no payment) for paid subscribers to get access. Then Perplexity does this 👎
RT Johnny Ho If your agent needs frontier web search, it's hard to beat Perplexity's Agent API (on any metric).
Introducing Web Search Benchmarks 🌐 Rankings of search tools across different models and configurations to help you decide how to ground your agent: https://openrouter.ai/benchmarks
RT Alex Atallah Great work by @perplexity_ai on these benchmarks! More info on each one, Pareto curves across models, and why we run them on https://openrouter.ai/benchmarks
Introducing Web Search Benchmarks 🌐 Rankings of search tools across different models and configurations to help you decide how to ground your agent: https://openrouter.ai/benchmarks
Perplexity's Search SDK, which makes Perplexity Computer the best-in-class product for wide and deep research, is now available to use inside any agentic harness!
Introducing the Perplexity Search SDK. It's an agent-first Python SDK that brings Perplexity's Search as Code approach to your applications. Agents can fan out multiple searches, then filter, dedupe, and rank results in code.
View quoted postPerplexity 👑
Introducing Web Search Benchmarks 🌐 Rankings of search tools across different models and configurations to help you decide how to ground your agent: https://openrouter.ai/benchmarks
Impressive numbers for a 700b parameter model!
Introducing GLM-5.3: Built to Code. Ready for Cyber Defense. - Top-tier coding and agentic capabilities, achieved through post-training on the 743B base model - A major leap in cybersecurity, setting a new standard among open models Tech Blog: https://z.ai/blog/glm-5.3
Perplexity Agent API is the best for web search and browsing agents
Sonar is moving to the Agent API. The Perplexity Agent API keeps grounded web search, and adds multi-step research, code execution, built-in tools, and access to multiple models through one API. On BrowseComp and WideSearch, Agent API more than doubles the best Sonar score.
RT Elon Musk Grok
Congrats to @SpaceXAI on one more amazing model: Grok 4.6. We benchmarked it as an orchestrator on our Wide-And-Deep-Research benchmark using the Perplexity Computer harness, and it neatly sits on the Pareto frontier of performance vs cost. Available to all Pro and Max users on
View quoted postCongrats to @SpaceXAI on one more amazing model: Grok 4.6. We benchmarked it as an orchestrator on our Wide-And-Deep-Research benchmark using the Perplexity Computer harness, and it neatly sits on the Pareto frontier of performance vs cost. Available to all Pro and Max users on Perplexity!
Grok 4.6 is now available in Perplexity and Perplexity Computer. On WANDR, it sits on the Pareto frontier of performance and efficiency, matching Fable 5 results at over 60% lower cost.
Gemini Flash models are great for fast, cost-efficient subagents inside any multi-model harness. We use them a lot inside the Perplexity Computer harness.
Our Flash models are workhorses that offer performance at a great price. So, we're shipping updates fast to get them in developers’ hands. And now just 3 weeks after launching 3.6 Flash, 3.7 Flash shows significant gains including in coding and agentic work. It delivers
View quoted postNemotron 3.5 Lightning available to all developers on the Perplexity Agent API. Input: $0.0115 per 1M tokens Output: $0.17 per 1M tokens
.@NVIDIA Nemotron 3.5 Lightning is now available in the Perplexity Agent API 🎉 An open 30B MoE model with 3B active parameters, built for the high-volume execution layer of always-on agents: tool calls, validation, and subagent work. Pair Nemotron Lightning 3.5 with frontier
View quoted postA great American open weights MoE model that can run efficiently on your laptop or local hardware like the DGX Spark! You can use the larger Nemotron Ultra on Perplexity!
Introducing NVIDIA Nemotron 3.5 Lightning⚡ An open 30B MoE model with 3B active parameters, built for always-on agents to complete high-volume, specialized tasks faster. It delivers up to 4x the output speed of similar-sized models.
RT NVIDIA AI Introducing NVIDIA Nemotron 3.5 Lightning⚡ An open 30B MoE model with 3B active parameters, built for always-on agents to complete high-volume, specialized tasks faster. It delivers up to 4x the output speed of similar-sized models.
Compute is the currency
enjoy US hosted K3 on perplexity agent api!
Access the power of @Kimi_Moonshot K3 in the Perplexity Agent API. Hosted exclusively on U.S.-based servers. Try it out today https://bit.ly/pplx-kimi-k3
View quoted postRT Paul Copplestone - e/postgres Supabase is now available on Perplexity Computer 🤖 From a Perplexity chat you can query your production data, look up users, and do basically anything on the Supabase platform @supabase 🤝 @perplexity_ai
RT Milton Friedman Quotes Milton Friedman on 4 ways to spend money: 1) Your money on yourself (you’re careful about both cost and quality) 2) Your money on others (you care about cost, less about quality) 3) Someone else’s money on yourself (you care about quality, not cost) 4) Someone else’s money on others (you care about neither) The last one is how government spending works.
Perplexity Computer’s goal is to provide the maximum intelligence for minimum cost with multi model orchestration. We’ve tested GPT 5.6 Terra extensively and found it to be a pretty good model: capable and cost-effective at the same time. We’re making it the default for all subagents inside Perplexity Computer harness. And are also offering it as an orchestrator model for all Computer users. Have fun!
GPT 5.6 Terra and Luna are now live in Perplexity Computer. Terra is the new default model for all Computer subagents, while Luna will serve as the primary model for scheduled automations. Terra is also available as an orchestrator model in Computer.
RT AleXandra Merz 🇺🇲 Today, I am using this transcript site: https://www.perplexity.ai/finance/SPCX/earnings?eventId=692542&tab=transcript
RT Wall St Engine If anyone wants to listen to the $SPCX and $AMD earnings call, you can find it here on Perplexity: 16:30: https://www.perplexity.ai/finance/SPCX/earnings?eventId=692542&tab=transcript 17:00: https://www.perplexity.ai/finance/amd/earnings?eventId=660937&tab=transcript
Verification is key when using agents for high stakes research!
Every numerical value in Perplexity Computer finance queries is traceable back to original values, with full calculation trace shown.
RT Jensen Huang Today, we’re launching Alpamayo 2 Super, our frontier open reasoning model for autonomous vehicles. Beyond seeing, Alpamayo understands and reasons through the complex world - thinks before it acts. It’s a powerful backbone for robotaxis, trucks, shuttles, delivery vans, tractors and the long tail of mobile robots—billions of autonomous machines someday. We’re releasing it for commercial use under OpenMDW-1.1 so teams can inspect it, fine-tune it and deploy it—open models advance safety and security. The next wave of AI is robotics—and it starts with autonomous vehicles. Great work, Alpamayo team! https://blogs.nvidia.com/blog/alpamayo-2-super-open-model-now-available
Two orders of magnitude improvements are quite rare. This is a big deal.
DeepSeek V4-Flash isn’t just cheaper per token. It reportedly completes the same benchmark tasks as Fable 5 at 105× lower total cost, according to @ArtificialAnlys ! That's precisely why the Flash release is, for me, the DeepSeek 2.0 moment. It will cause a huge stir.
RT Perplexity Developers The Perplexity remote MCP server is live. You can now connect Perplexity to Claude Code, Cursor, or VS Code with just your API key. Nothing to install. https://docs.perplexity.ai/docs/getting-started/integrations/mcp-server#remote-mcp-server
RT DeepSeek 🚀 DeepSeek-V4-Flash Official API is now LIVE in public beta! 🔷 We’ve massively upgraded its Agent capabilities—benchmark scores are now far surpassing the V4-Pro-Preview. Check out the massive performance leap below! 👇 🔷 The official V4-Flash now natively supports the Responses API format and is fully adapted for Codex! Check out the configuration details in our official API docs: https://api-docs.deepseek.com/quick_start/agent_integrations/codex