Realtime AI News
Perplexity brings its local Portable Computer agent to Windows PCs with RTX GPUs
Perplexity made its on-device Portable Computer agent available in the Windows app on September 14, running entirely on NVIDIA GeForce RTX or RTX PRO GPUs with at least 24GB of VRAM and open to Pro and Max subscribers. Unlike the cloud version, the model, agent harness, orchestrator and scheduler all run on the machine, so sensitive data stays on the device and locally completed work does not consume Perplexity Computer credits.
Perplexity brought Portable Computer to Windows on September 14. The local agent had previously been aimed at NVIDIA DGX Spark systems and launched first on Linux, but it is now inside the Windows Perplexity app, letting users run a cloud-free AI agent on an ordinary Windows machine.
The trade-off is an explicit hardware bar: an NVIDIA GeForce RTX or RTX PRO GPU with at least 24GB of VRAM. NVIDIA said support for its DGX Station workstation is expected soon, which keeps the current addressable hardware concentrated among high-end laptops, mobile workstations and desktops with large amounts of video memory.
The bigger difference from cloud agents is where the work happens. Portable Computer runs the model, the agent harness, the orchestrator and the scheduler entirely on the device. Perplexity made the same point when it first shipped the product in August: locally completed work does not consume Perplexity Computer credits, and sensitive information does not have to leave the machine.
The feature is available to Pro and Max subscribers across individual and enterprise plans through the existing Perplexity app for Windows, which users download from the Microsoft Store. After picking a local model from the dropdown, they can pull it down in one click instead of researching models and configuring a local inference stack themselves.
To make the agent useful, Portable Computer ships connectors for Outlook, OneDrive, Word, Google Drive, Gmail, Slack and GitHub, and it can reach other desktop applications through local Model Context Protocol servers running on the Windows device, so the local model can use those tools and data inside automated workflows.
Perplexity offered a logistics example: a controller schedules the agent each morning to reconcile freight invoices against the carrier's local rate sheets, flag duplicate charges or rates that deviate from contract, and generate an exception queue. When a task needs current information or stronger reasoning, the agent can call Perplexity Search or one of more than fifteen frontier models after asking the user's permission.
Why it matters: local agents address two of the biggest pain points of cloud agents at once, since data never leaves the device and repetitive on-device work no longer burns credits. It also shows that competition among agents is shifting from raw cloud model capability toward on-device efficiency and how deeply the agent plugs into everyday office software.
What to watch next: which additional local models enter the picker, when workstation platforms such as DGX Station are formally supported, and whether enterprise buyers lean toward this nearly fully local deployment path to satisfy data-compliance requirements.
Why it matters
Perplexity is now selling its agent on data sovereignty and cost control rather than model capability alone, positioning local inference as a direct alternative to cloud agents. The 24GB VRAM floor confines near-term adoption to high-end RTX hardware, which points to professionals and enterprise deployments before mainstream users.
Nearby Updates
All09/15, 18:41
HiDream.ai releases HiDream-O1-Video-1.0, a natively omni-modal video model that lands in the global top tier
HiDream.ai released HiDream-O1-Video-1.0 on September 15, describing it as the first natively omni-modal video generation model, with text, image and video inputs producing 1080p clips of five to twenty seconds. It ranked fourth on the Artificial Analysis Image to Video Leaderboard (With Audio) and eighth on Arena.ai's image-to-video blind evaluation, while the company announced a C+ round backed by Newmicro Capital, Jiaozi Capital and ICBC Capital.
09/15, 17:58
MoleculeMind pushes QuantaMind to 100,000-atom reaction simulations on a single GPU
MoleculeMind says its reactive machine-learning force field QuantaMind can now simulate reactions in 100,000-atom systems for hundreds of nanoseconds at near quantum-chemistry accuracy, at about 0.25 seconds per step on a single GPU. The underlying Science Advances paper ran a continuous 6-nanosecond simulation of a 17,792-atom PETase system, covering proton transfer, bond breaking and formation and the full catalytic cycle, with agreement above 0.99 against quantum-mechanical checks.
09/15, 20:00
Salesforce and Nvidia Launch Koa, a Reasoning Model Trained for Sales and Support
TechCrunch reports that Salesforce and Nvidia have launched Koa, a reasoning model built on Nvidia's open-weight Nemotron and trained for sales, marketing, and customer-support tasks. The pairing signals that enterprise software vendors can now train vertical reasoning models on open weights instead of depending on frontier labs.
09/15, 20:03
The Information: ByteDance first-half profit drops to $20 billion as AI spending weighs
The Information reports that ByteDance's first-half profit fell to 20 billion dollars, with heavy spending on artificial intelligence cited as the main drag. The public summary does not break out revenue mix or capital expenditure, but it offers a concrete signal of how far AI investment is now compressing profit at a leading platform.