You can now run Harbor evals on Vercel Sandbox. Harbor is the open-source harness behind Terminal-Bench , whose registry includes many other benchmarks such as SWE-bench, tau3-bench and OSWorld. Pass --env vercel to harbor run and each tria…
Firehose
Filtered to Companies · clear filters
Browse: People · Companies · Papers · Podcasts · Hacker News · Deep dives
Browse by tag
[email protected] adds Notion skills databases as an install source for agent skills . Notion skills are reusable agent skills written as Notion pages. Teams author, review, and update them in the workspace they already use, then install them in…
Developers, a new layer called Omnigent in Databricks enables engineers to define an agent once, including the model, tools, policies, and limits, and run it across any harness, reducing the need to rebuild and manage multiple instances. Omnigent integrates with the Foundation Model APIs for unified cost and governance tracking. Additionally, a new web search component called Nimble, which can adapt to specific use cases and self-learn the best retrieval methods, can be integrated to improve the accuracy and efficiency of web search. AI summary
AIOps (Artificial Intelligence for IT Operations) combines AI and machine learning with observability to automate IT operations, detecting anomalies, correlating events, identifying root causes, and automating incident response, ultimately reducing downtime risk, cutting alert fatigue, and accelerating decision speed. AI summary
A rewrite this size wasn't affordable before agents. Here's what porting the Copilot agent runtime to 800,000 lines of production Rust actually took. The post Migrating the GitHub Copilot runtime to Rust, using Copilot appeared first on The…
You can now connect native Marketplace resources to custom environments . Previously, resource connections could only target production, preview, and development environments. Choose custom environments when connecting a resource from the V…
Anthropic has introduced the Life Sciences Verification Program (LSVP), a beta program offering refined safeguards for biology-related work, allowing life science professionals to access Anthropic's Mythos, Opus, and Sonnet models. LSVP grants are available for teams and institutions, with a verification process reviewing research credentials, security standards, and ethical research oversight. The program's safeguards aim to protect against access compromise, insider threats, and agent misuse, with monitoring usage against intended use cases and data retention for 30 days to identify potential misuse. AI summary
A modern storefront can look healthy while malicious JavaScript quietly siphons revenue, hijacks clicks, or rewrites analytics. See how Cloudflare's machine learning models surface evasive client-side attacks for analyst investigation.
Hobby projects now retain fewer deployments past the 30-day retention window. Hobby teams get 10GB of Deployment Storage . Every deployment you keep uses some of it, and going over the limit can block you from deploying until you free some …
MATCH_RECOGNIZE is a new SQL operator available in Public Preview that allows detecting patterns and sequences from event data using regex-like pattern-matching, simplifying pattern detection and sequence analysis across various industries. AI summary
OpenAI shares a framework for tracking, investigating, and disclosing model misalignment, alongside six reports of unexpected or concerning model behavior.
Mem0 is now available as a native integration on the Vercel Marketplace , giving your AI agents and apps long-term memory. Mem0 remembers user preferences, facts, and context across sessions, so your app stops starting from scratch. Install…
Builds using Secure Compute or Static IPs now start 64% faster, with the average time from deployment creation to build start dropping from 6.7 seconds to 2.4 seconds. Previously, each build waited for a new build container to boot with its…
OpenAI and AARP are bringing free, hands-on ChatGPT workshops to 1,000 older adults across 10 U.S. cities to build practical AI skills safely.
Explore new AI-powered advertising experiences from OpenAI, including Sponsored Agents, tools for marketers, and integrations with HubSpot and Shopify.
Learn how ChatGPT Work and Codex analytics help teams understand AI usage and spend, identify training needs, and connect adoption to business outcomes.
Open, private and multilingual AI is coming to your web browser. Mistral and Mozilla team up to put powerful, trustworthy AI where you already browse.
New OpenAI Economic Research shows how workers use AI beyond traditional roles and which new activities become recurring parts of their work.
Jev from TypeSafe AI is now available on AI Gateway . Jev is a probabilistic decision model for software: state goes in, typed Choice, Score, and Boolean answers come out. Regular language models generate text one token at a time, which the…
Is Agentic reports now let you view your checks through one of four site types: Docs & content, Business, App, or Commerce. For example, the Commerce view highlights payment and checkout standards like x402, UCP, and ACP, while the App view…
Google has launched Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, its most advanced live dialogue models yet, offering more intuitive and intelligent conversations. These models handle complex reasoning, real-time visual context, and background task execution without interrupting the conversation. AI summary
Databricks' Genie and AI business processes can operationalize ML insights into governed business operations for energy theft detection, enabling teams to accelerate the loop from flagged meter to recovered revenue to a safer household. A governed workflow connects interpretation, investigation prioritization, dispatch-ready reporting, and recovery workflows, with Lakebase maintaining live case state and recovery totals. Trusted answers and metrics are provided through Genie One, Unity Catalog, and Unity Gateway, enabling leaders to ask questions in plain English and receive answers grounded in governed data. AI summary
Researchers at IBM developed a method to improve the consistency of large language models (LLMs) like GPT-4.1, which can significantly impact their reliability in mission-critical applications. By analyzing an agent's past trajectories and identifying "flat" decisions, where the model is uncertain, they created a new type of guideline that helps stabilize these decisions. This approach, called consistency guidelines, can improve the Pass^5 metric, which measures the fraction of tasks an agent succeeds on all runs, by up to 22.9 percentage points. AI summary
We’re moving beyond traditional text translation to build models that understand the world’s rich, living languages exactly as they are expressed.
The true measure of AI is who it helps. Here’s how it’s impacting lives today. We're focused on key areas where advanced technology can help make extraordinary progress …
Explore this collection to see how experts and local leaders are using AI breakthroughs to ensure everyone can share the opportunity of AI.
You can now scope access to individual Workers and assign narrower Developer Platform roles, so teammates, CI tokens, and agents get only the access they need to debug, deploy, or monitor safely.
Cloudflare is giving site owners a way to stay discoverable while disallowing AI training. New controls and an Accountable designation establish a shared model with Apple, Google, and Microsoft.
We’ve translated ATLAS’s millions of global data points into an interactive, open-access experience.
Delphi on Vercel 10 engineers with no dedicated infrastructure role Everyone ships code, including product and design 100+ production deploys a day behind feature flags Delphi builds digital minds. They capture what someone has written, rec…
Databricks' marketing team uses Genie, an AI analytics assistant, 3x more often in decision-making, with over 85% adoption across the marketing organization. They achieved this by building a governed Marketing Lakehouse, documenting data and business context, encoding verified answers and examples, teaching Genie the language of their business, and continuously evaluating and improving the system through user feedback. AI summary
Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking from Google are now available on AI Gateway. Both models support real-time spoken interactions for voice assistants, conversational experiences, and applications that respond through aud…
Managed Postgres should take routine database operations such as patching, scaling, failover, and backups off the database team's plate. Lakebase automates these operations on serverless infrastructure with automatic scaling, scale-to-zero, point-in-time recovery, branching, pgvector, and PostGIS. AI summary
The AI SDK harness layer now supports authenticating harnesses through their native subscriptions, where the underlying harness supports them. The harness layer runs different coding agents through the same HarnessAgent interface, so you ca…
Databricks now supports on-demand state repartitioning for Apache Spark Structured Streaming, allowing users to resize partitions without rebuilding checkpoint state, enabling more flexible tuning and scaling of stateful streaming queries. This feature is available in Databricks Runtime 18 and above with the RocksDB state store provider. AI summary
Christina Koch sits down with James Manyika, Google’s Senior Vice President of Research, Labs, Technology & Society.
DevFest 2026 is back and here’s how you can connect with one of the more than 800 global events to build, secure, and scale in the agentic AI era.
Fyxer uses OpenAI models, fine-tuning, memory, and real user feedback to organize inboxes and draft emails in each user’s voice.
Perplexity uses Astra to write communications, change software, and monitor production systems, and checks in much less frequently than with earlier models.
Conversational AI can help payer finance leaders decompose variances in medical loss ratio (MLR) across claims, utilization, cost, and population risk in minutes, without waiting on analysts or reports. However, AI needs payer-specific context to earn trust, and without it, AI amplifies confusion instead of resolving it. A unified payer intelligence foundation, combining governed data and AI capabilities with payer-specific data and business context, can help finance leaders understand what caused the variance and take corrective action. AI summary
If you can write down how you do your work, you can automate it. Here's what I did to support GitHub's APAC marketing team. The post Marketing ops as code: Automating events from planning to follow-up on GitHub appeared first on The GitHub …
Lakeflow Connect provides native, fully managed connectors for various SaaS applications, databases, and file sources, ingesting data directly into Databricks Platform. These connectors can be set up via a point-and-click UI or a simple API, and the ingested data is incrementally ingested as governed, managed tables in Unity Catalog. AI summary
Cognition's autonomous software engineer Devin uses GPT-6 Astra to test its own work, generating recordings and reports that help engineers review less code and ship more. This AI-powered testing tool improves Devin's ability to test software and show results, enabling faster bug fixes and more efficient code review. By leveraging GPT-6 Astra, Cognition aims to reduce manual code review and increase shipping efficiency. AI summary
Introducing ChatGPT Images 2.5, Linguistic drift at the frontier, a paper on Random Attention: Rethinking KV Cache Eviction for Efficient Reasoning, and many more!
Cloudflare CASB policies introduce a native automation engine built directly on the Cloudflare developer platform to remediate SaaS risks automatically. Security teams can now design event-driven logic to revoke risky file shares and send w…
Learn how OpenAI evolved Habitat from a Python library into a globally distributed storage platform serving 1 billion ChatGPT users and 22M requests per second.
Every Vercel Sandbox now includes 64 GB of storage, up from 32 GB. This includes sandboxes created from a Vercel Managed Image or custom image, as well as those configured with the deprecated runtime property. The additional space provides …
Tailscale on Vercel Hundreds of AI models shipped to customers in-product Model access granted and revoked by tailnet network identity Went from model routing prototype to paying customers in months Tailscale connects a company's laptops, s…
Featured on Vercel 3 engineers supporting 3 brands and 100,000+ users on Vercel Migrated 374 Sanity sites from AWS Elastic Beanstalk to Vercel AI SDK and AI Gateway power Featured's chat bot across 17 models Workflow SDK replaced custom lon…
Pro and Enterprise teams can now control who can create and manage Vercel Connect connectors. Connectors let applications and agents access external services using credentials managed by your team. Owners can enable this restriction under C…
Checking agent-generated code usually means hopping between tabs. Learn how to view diffs, run terminal commands, and preview web apps side by side in the GitHub Copilot app. The post GitHub Copilot app for Beginners: Using the diff, termin…
The AI SDK harness layer now supports GitHub Copilot through the official @ai-sdk/harness-github-copilot adapter. The harness layer lets your application run different coding agents through the same HarnessAgent interface, so you can switch…
FastAPI frontends and static files, served with app.frontend() or StaticFiles , are now promoted to the Vercel CDN at build time. Requests for those paths are served directly from the CDN, without invoking your Vercel Function. FastAPI eval…
Search can help runners get race-day ready with registration alerts, tailored training plans, and more.
Researchers use Codex and ChatGPT to accelerate the search for new antimicrobial molecules by leveraging AI to analyze vast genome and protein datasets, identify patterns, and prioritize candidate molecules for experimental testing. The lab's approach combines deep-learning models with human expertise from multiple scientific disciplines to decipher the organizing principles of life that give rise to functional molecules. Codex and ChatGPT are used to brainstorm hypotheses, write code, process datasets, and analyze results, helping to bridge gaps between scientific disciplines and accelerate the discovery process. AI summary
Read restrictions and catalog labels from Apache Iceberg™ standardize delegated enforcement and governance context portability across engines and catalogs, respectively, addressing fragmented enforcement and enabling unified governance across the Open Lakehouse. AI summary
OpenAI has introduced the Data agent in ChatGPT Work, enabling users to turn questions into analysis and action by connecting to approved data sources, including Amazon Redshift, Datadog, Google BigQuery, and more. The agent uses business terms, metric definitions, and data relationships to interpret the data and build interactive dashboards. Users can interact with the data in plain language, refining the analysis in one conversation without writing queries or learning a new analytics tool. AI summary
Databricks has improved the Lakebase Postgres compute cache, increasing throughput by up to 2x and reducing latency. The new cache path uses larger shared buffers, which can hold up to 75% of available DRAM, and enables autoscaling. This reduces double buffering and improves cache hit rates, resulting in lower CPU usage and faster access to data. AI summary
1.1.1.1 now validates DNSSEC signatures using NIST’s post-quantum ML-DSA-44 algorithm. Here is how we manage 2,420-byte signatures and downgrade risks at scale.
Cloudera and Mistral join forces to bring specialized, sovereign AI intelligence to enterprise data, helping regulated industries innovate on their own terms.
OpenAI has partnered with the U.S. General Services Administration (GSA) to provide free access to its AI tools, including GPT-6 Astra, for federal, state, local, and tribal governments, with 50% off usage costs. The agreement also expands support for public-sector cyber defenders and provides training and enablement to help government leaders and cyber defenders use the model effectively. Eligible entities will have access to Daybreak Blue, a model that can help identify vulnerabilities and accelerate development of critical systems. AI summary
OpenAI has introduced ChatGPT for Financial Services, a tailored ChatGPT Work experience that combines built-in financial data with GPT-6 Astra's reasoning capabilities to help teams develop research, financial models, and customized client materials. The product includes premium financial data from providers like Daloopa, PitchBook, and LSEG News, and offers state-of-the-art frontier intelligence, including GPT-6 Astra. AI summary
In August, we experienced five incidents that resulted in degraded performance across GitHub services. The post GitHub availability report: August 2026 appeared first on The GitHub Blog .
Async GRPO with LoRA across HF Jobs uses a bucket, a proxy, and no NCCL, leveraging Hugging Face Jobs and Storage Buckets to train a LoRA adapter and sync only that adapter to vLLM replicas. The proxy routes each rollout to the replica that already holds its KV prefix and broadcasts adapter loads to all replicas. AI summary
You can now build and deploy long-running, tool-using agents with the OpenAI Agents API on Vercel. OpenAI manages the agent loop and session state, while Vercel hosts the application and connects each session to Vercel Sandbox for code exec…
Build and launch cloud agents with the Agents API, a managed service powered by the Codex harness for orchestration, long-running sessions, and tool use.
Vercel Sandbox can now run in all 20 Vercel compute regions , up from four. Running sandboxes closer to the databases, storage, and other services they access reduces latency. Teams can also keep sandbox workloads in approved regions to sup…
Tako Search is free exclusively on AI Gateway through September 30th. Tako Search is a model-agnostic search tool that works with any model on AI Gateway. You can switch models while keeping the same search integration, without a separate T…
Every request to Vercel passes through our CDN , which executes on average over 80 million routing instructions per second. Part of that work is looking up metadata to determine which paths exist and how to serve them. When that metadata is…
Gradio's Workflow1111 is a single canvas that integrates various media pipelines, including text-to-image, hi-resolution fix, image-to-image, prompt matrix, VLM interrogate, detection to inpaint masks, and image-to-video, using SOTA models and operator types like fn, model, and space nodes, allowing for free parallelism and zero-code REST/MCP endpoints. AI summary
GPT-Live-1 in the API introduces a full-duplex conversation model that enables developers to build more natural voice experiences with improved interruption handling, reasoning and tool calling delegation, tone, pace, and style customization, and silent context management. AI summary
Here is a summary of the article in 3 plain sentences for a developer/AI-ML audience: Financial services leaders are asking about the governance and cost implications of implementing AI solutions, with questions including how to ensure AI systems are trustworthy, how to use real-time data without adding complexity, and how to manage AI costs as usage expands. To address these concerns, Databricks is showcasing its platform and AI capabilities at Sibos 2026, including its ability to provide governed data and AI for financial services workflows and its solutions for real-time data and AI governance. The company is also hosting executive meetings and providing demos at the event, which takes place September 28-October 1 in Miami, Florida. AI summary
Paul Christiano has joined the OpenAI Foundation Board as a non-voting observer on the OpenAI Group PBC Board, bringing his experience in government and AI alignment to the organization, and will also join the Safety and Security Committee. AI summary
Databricks provides a unified platform for end-to-end Solvency II reporting, connecting data ingestion, quality checks, reserving, capital calculation, quantitative reporting templates, own risk and solvency assessment, governance, approvals, and disclosure. This approach automates ingestion checks, provides a single control view, and supports governance, audit trails, and AI-assisted review. AI summary
Discover how filmmakers and Google DeepMind used AI to recreate a couple's unrecorded past in the short film "Love, Rendered."
Track live game feeds, explore detailed stats, and get custom fantasy recommendations directly in Search this season.
IBM has released the top-performing zero-shot time series forecasting model, Granite Time Series PatchTST-FM-r2, with a commercial-friendly license under Apache 2.0 and OpenMDW 1.0. The model achieves strong zero-shot performance on the GIFT-Eval benchmark, ranking #2 overall among replicable, zero-shot models, and outperforms pretrained models despite being smaller in size. AI summary