Sign in

Salesforce AI Research

@sfresearch.bsky.social
166 followers 57 following 215 posts

We advance state-of-the-art #AI techniques paving the path for innovative products at @Salesforce.com. Focus areas: #AIAgents, #EnterpriseAI, #EGI, and #TrustedAI.

PostsRepliesMedia
Salesforce AI Research @sfresearch.bsky.social · 16h
(1/5) We are pleased to announce our participation in COLM 2026, the Third Annual Conference on Language Modeling, at the Hilton Union Square in San Francisco, October 6–9. Our researchers will present 4 accepted papers. Full list below ⬇️ #COLM2026
100
Salesforce AI Research @sfresearch.bsky.social · 08/09/2026
🧵 (1/3) We're heading to #ECCV2026 in Malmö, Sweden. 🇸🇪 This week, our team will present two new papers spanning physical generative reasoning and video question answering. More below ⬇️
121
Salesforce AI Research @sfresearch.bsky.social · 02/09/2026
(1/13) 🎉 We are pleased to announce our participation in #EMNLP2026, the Conference on Empirical Methods in Natural Language Processing, Budapest, Hungary, October 24–29. 🇭🇺 Our researchers will present 12 accepted papers spanning agentic reasoning, evaluation, and multimodal AI. Full list below ⬇️
100
Salesforce AI Research @sfresearch.bsky.social · 02/09/2026
Traditional monitoring tells you what happened last week; #OperationalIntelligence tells you what's happening now and why. Our EVP & Chief Scientist Silvio Savarese breaks down the shift to continuous, proactive systems that surface what matters before it becomes a problem. sforce.co/4cquhTH
sforce.co
Operational Intelligence: Turning Enterprise Data Into Enterprise Decisions with AI
Silvio Savarese on operational intelligence: the AI layer that makes enterprise signal legible, so leaders can act before the moment passes.
110
Salesforce AI Research @sfresearch.bsky.social · 10/07/2026
What happens when AI agents start negotiating with each other, and who's responsible when they do? Silvio Savarese and Sabastian Niles break down the New AI Trust Architecture: 5 requirements for safe, accountable agent-to-agent communication. sforce.co/4gwgT3p #AgenticAI #AItrust #AIgovernance
salesforce.com
The New AI Trust Architecture: 5 Requirements for Agent-to-Agent Communication
5 requirements for trustworthy AI agent-to-agent communication, from Salesforce's Chief Scientist and Chief Legal Officer. Standards, identity, and accountability, explained.
010
Salesforce AI Research @sfresearch.bsky.social · 09/06/2026
(1/5) Model cards are nutrition labels for AI. Now they include environmental impact. 🌱 @Salesforce.com is adding standardized energy + carbon metrics: sforce.co/4umu8qm
sforce.co
Measuring AI’s Environmental Impact: How We’re Operationalizing Transparency Through Model Cards
Today, Salesforce is expanding its AI model cards with standardized environmental impact metrics. This update helps customers better understand the energy
100
Salesforce AI Research @sfresearch.bsky.social · 03/06/2026
(1/8) Can Language Models Remember What They Learn? LLMs learn from feedback. But most post-training is amnesiac: rollout → reward → update → forget. What if you keep the signal? Procedural Memory Distillation (PMD): learning from experience, not just feedback. 🧵
110
Salesforce AI Research @sfresearch.bsky.social · 02/06/2026
6th Multimodal Algorithmic Reasoning Workshop at #CVPR2026 is Thursday (6/4), 8:55 AM–12:30 PM MDT, Room 601, Colorado Convention Center. 👥 Keynote by Juan Carlos Niebles @jcniebles.bsky.social, organized by Honglu Zhou @hongluzhou.bsky.social 📆 sforce.co/4ueOT7j
011
Salesforce AI Research @sfresearch.bsky.social · 28/05/2026
Can Language Models Remember What They Learn? Introducing Procedural Memory Distillation (PMD): sforce.co/4dAjQOu PMD turns model attempts into reusable training memory, conditions a self-teacher on it, and distills the guidance into the student's weights.
sforce.co
Can Language Models Remember What They Learn?
Post-training methods (RLVR, On-policy distillation) are Episode-local Language models are getting better at learning from feedback during post-training. In reinforcement learning with verifiable rewa...
020
Salesforce AI Research @sfresearch.bsky.social · 26/05/2026
📣 1/5 Counterparty Modeling is Not Strategy: The Limits of LLM Negotiators sforce.co/3RTTU7Q New research tests whether LLM agents can turn knowledge of a partner's preferences into better deals. The finding? LLMs can know what the other side wants without bargaining strategically on it.
sforce.co
Counterparty Modeling is Not Strategy: The Limits of LLM Negotiators
Negotiation requires more than inferring what the other side wants: it requires using that information to make advantageous offers and counteroffers over multiple turns. We study whether large languag...
112
Salesforce AI Research @sfresearch.bsky.social · 22/05/2026
1/5 RLVR trains LLMs with pass/fail rewards — but every near-miss rollout is wasted. What if models could actually *learn* from their mistakes? New paper: "Learning from Language Feedback via Variational Policy Distillation" Read: sforce.co/4uv2f0k 🧵👇
111
Salesforce AI Research @sfresearch.bsky.social · 20/05/2026
When AI Becomes Invisible: The Rise of Ambient Intelligence Silvio Savarese on the shift from AI as a tool you go to, to a presence already there: always-on, aware, adaptive, and anticipatory. sforce.co/3Pwo68s #FutureOfAI #EnterpriseAI #AmbientIntelligence
sforce.co
When AI Becomes Invisible: The Rise of Ambient Intelligence
A perspective on the future of enterprise AI that is ambient: context-aware, always on, and invisible to users.
000
Salesforce AI Research @sfresearch.bsky.social · 18/05/2026
Nvidia has adopted our xLAM function calling dataset in NeMo Gym, their new library for building RL training environments for LLM agents. Check out our function calling dataset: sforce.co/4uZkL0P And see how it works in Nemo Gym: bit.ly/4tuu0pg #AgenticAI #ReinforcementLearning
010
Salesforce AI Research @sfresearch.bsky.social · 14/05/2026
SFR-VibeTrain: The Agent That Trains Agents A messaging-native agentic control plane for model and agent training. Describe the task, upload data (or don't), and it reasons, launches, analyzes, ablates, and deploys. sforce.co/4dwzYzo #FutureOfAI #EnterpriseAI #AgenticAI
sforce.co
SFR-VibeTrain: The Agent That Trains Agents
What if launching an RL training run felt less like operating a GPU cluster and more like talking to a sharp research engineer in Slack? Training AI models is still strangely artisanal, involving…
010
Salesforce AI Research @sfresearch.bsky.social · 08/05/2026
A great moment for the lab! 🎉 "Don't Stop Early" earned its spot on the Industry Track for advancing enterprise deep research at scale. Controlled information flow and evidence-aware termination prevent premature stopping and uneven coverage. bit.ly/49fk2zQ #ACL2026
bit.ly
Dont Stop Early: Scalable Enterprise Deep Research with Controlled Information Flow and Evidence-Aware Termination
Enterprise deep research often fails to produce decision-ready reports due to uneven information coverage, context explosion, and premature stopping. We propose a scalable Enterprise Deep Research (ED...
130
Salesforce AI Research @sfresearch.bsky.social · 06/05/2026
🎉 3 papers from @salesforce.com AI Research accepted to #ICML2026, spanning Mixture-of-Experts efficiency and multi-agent reasoning + verification. Learn more ↓
100
Salesforce AI Research @sfresearch.bsky.social · 01/05/2026
Our Infrastructure & DevOps team was featured by Google Cloud. Lavanya Karanam and Avinash Gudagi on how @salesforce.com is solving one of the hardest problems in large-scale AI training: keeping powerful GPUs fully fed. bit.ly/4n6HDc3 #FutureOfAI #EnterpriseAI
000
Salesforce AI Research @sfresearch.bsky.social · 30/04/2026
3 AI trends shaping #EnterpriseAI: simulation environments, ambient intelligence, and agent-to-agent communication. Itai Asseo on our focus on Enterprise General Intelligence (#EGI) in @gregorojstersek.bsky.social's #TDX26 recap. bit.ly/4mT4irW #FutureOfAI
bit.ly
3 Key AI Trends and How Salesforce Engineers use AI
Recap from the Salesforce TDX 2026: AI Agents, 3 key AI trends and how engineers use AI.
000
Salesforce AI Research @sfresearch.bsky.social · 29/04/2026
(1/5) 📊 Don't Stop Early: Scalable Enterprise Deep Research with Controlled Information Flow and Evidence-Aware Termination bit.ly/49fk2zQ
100
Salesforce AI Research @sfresearch.bsky.social · 24/04/2026
Congratulations to our researches on winning an Outstanding Paper Award at #ICLR2026 for LLMs Get Lost in Multi-Turn Conversation! 🏆 Authors: Philippe Laban, Hiroaki Hayashi, Yingbo Zhou, Jennifer Neville bit.ly/3ZoC6T1
0101
Salesforce AI Research @sfresearch.bsky.social · 24/04/2026
Attending #ICLR2026? Meet us at @salesforce.com Booth #203 to hear the latest on how we're evolving enterprise general intelligence. For more, follow our journey at ICLR: bit.ly/4vLy9qohttps://bit.ly/4vLy9qo #FutureOfAI #EnterpriseAI
010
Salesforce AI Research @sfresearch.bsky.social · 23/04/2026
📍 Live from Rio: 21 papers at #ICLR2026, starting today. Our work spans agent architectures, reasoning, evaluation, deep research reliability, and scalable RL, tackling what matters most for #EnterpriseAI. sforce.co/4vRgqOs #FutureOfAI
sforce.co
Salesforce AI Research at ICLR 2026
Salesforce AI Research will present 21 accepted papers at ICLR 2026, the Fourteenth International Conference on Learning Representations. The conference runs April 23–27 at the Riocentro Convention an...
010
Salesforce AI Research @sfresearch.bsky.social · 22/04/2026
Can AI coding assistants maintain effectiveness as codebases scale 100×? LoCoBench-Agent evaluates agents across 10K to 1M token contexts, spanning 8,000 scenarios in 10 programming languages and four difficulty tiers. sforce.co/4txCDj4 👥: Jielin Qiu & Huan Wang #EnterpriseAI #AgenticAI
sforce.co
Beyond 100K Tokens: Evaluating AI Agents in Long-Context Software Engineering
As codebases grow to millions of lines of code, can AI agents still understand, reason, and code effectively? LoCoBench-Agent delivers the answer: a comprehensive benchmark for evaluating AI coding as...
010
Salesforce AI Research @sfresearch.bsky.social · 21/04/2026
The Illusion of Certainty: Decoupling Capability and Calibration in On-Policy Distillation: bit.ly/48iccVY On-policy distillation boosts accuracy but causes severe overconfidence. CaOPD uses a student-grounded empirical target for Pareto-optimal calibration. Code: bit.ly/4cUtCKO
010
Salesforce AI Research @sfresearch.bsky.social · 21/04/2026
At #TDX26, Itai Asseo @iiitaiii.bsky.social on Enterprise General Intelligence: refining generic LLMs into reliable enterprise agents through AI Foundry and Agentforce Labs, including eVerse, the Learning Engine, and Agent Startup. sforce.co/3OrRYCb #FutureOfAI #EnterpriseAI
sforce.co
TDX Live Blog: All the Highlights You Missed
On-the-ground reporting from the must-attend developer event for the Agentic Enterprise
010
Salesforce AI Research @sfresearch.bsky.social · 20/04/2026
(1/7) We have 6 papers accepted to ACL 2026, advancing work across web agent evaluation, LLM reasoning verification, uncertainty quantification, long-context efficiency, and multilingual judge systems. ACL 2026 takes place July 2-7 in San Diego, California. #ACL2026 #FutureOfAI #EnterpriseAI
121
Salesforce AI Research @sfresearch.bsky.social · 17/04/2026
That's a wrap on #TDX26! The @salesforce.com AI Research team was on the ground in San Francisco, connecting with the community and sharing the latest from our labs. From research demos to conversations about what's next in #EnterpriseAI, it was great to be part of the energy!
010
Salesforce AI Research @sfresearch.bsky.social · 16/04/2026
Reliability over raw power. 🔬 @aimagazine.bsky.social features Silvio Savarese on why enterprise AI value comes from integrated systems—not bigger models—and how AI Foundry is building that foundation. bit.ly/41z0G4E #SystemLevelAI #AIFoundry
bit.ly
Salesforce AI: Reliability Trumps Raw Model Capability
As AI matures, enterprise success hinges on integrated systems that deliver consistent performance across the most complex professional business workflows
010
Salesforce AI Research @sfresearch.bsky.social · 12/04/2026
AI Foundry: Turning foundational research into enterprise AI products faster. Silvio Savarese and Itai Asseo in TechFinitive: bit.ly/4ccYmFy #FutureOfAI #EnterpriseAI #AgenticAI
bit.ly
Salesforce launches AI Foundry as it doubles down on three "big bets" to accelerate enterprise AI
Details on Salesforce AI Foundry, launched today to help AI researchers, customers and partners to collaborate in a safe environment
010
Salesforce AI Research @sfresearch.bsky.social · 11/04/2026
The model wars are over. Enterprise AI success now lives at the system level. Silvio Savarese and Itai Asseo in @technologymag.bsky.social: bit.ly/4ttSBu4 #FutureOfAI #EnterpriseAI #AgenticAI
bit.ly
Salesforce AI Foundry: System Reliability Beats Model Power
The era of the model wars is over, with enterprise AI success occurring at the system level, demanding reliability and full integration over model power
010
Salesforce AI Research @sfresearch.bsky.social · 10/04/2026
Ambient intelligence is moving from research into live sales and service workflows. Silvio Savarese discusses with @cxtoday.com: bit.ly/47DUhJ1 #FutureOfAI #EnterpriseAI #AgenticAI
bit.ly
Salesforce Brings Ambient Intelligence to Sales Calls
Salesforce showcases ambient intelligence for real-time sales workflows and highlights Agentforce upgrades focused on scale, governance, and enterprise outcomes.
010
Salesforce AI Research @sfresearch.bsky.social · 09/04/2026
Three agentic AI trends shaping the enterprise through 2027. Silvio Savarese and Itai Asseo discuss AI Foundry with @cio.com: bit.ly/4sg6iMf #FutureOfAI #EnterpriseAI #AgenticAI
bit.ly
Salesforce AI Research identifies trends shaping agentic AI
Simulation environments, agent-to-agent ecosystems, and ambient intelligence will be at the heart of the Salesforce product roadmap through AI Foundry initiative.
020
Salesforce AI Research @sfresearch.bsky.social · 09/04/2026
One demo. Reliable replay. No cloud calls. GPA turns a single recorded workflow into deterministic desktop automation, entirely on-device. 🔎 Explore GPA: bit.ly/48r7Onp 📖 Read the blog: sforce.co/4sYdhu8 #EnterpriseAI #GUIAutomation
020
Salesforce AI Research @sfresearch.bsky.social · 08/04/2026
Why #EnterpriseAI demands a shift from models to systems. Itai Asseo discusses AI Foundry with @diginomica.com: bit.ly/4scZwGT #FutureOfAI #AgenticAI
bit.ly
The big bets are on as Salesforce pitches the need for enterprise transition from model to system level AI
Itai Asseo, VP of Salesforce AI Research, explains some new enterprise realities on the way.
021
Salesforce AI Research @sfresearch.bsky.social · 07/04/2026
From One Demo to Reliable Automation: How GPA Reimagines GUI Process Automation sforce.co/4sYdhu8 Show it a workflow once. GPA replays it reliably, locally, and without brittle scripts to maintain. #FutureOfAI #EnterpriseAI #AgenticAI #GUIAutomation
sforce.co
From One Demo to Reliable Automation: How GPA Reimagines GUI Process Automation
Tired of GUI automation that breaks after one demo? Discover how GPA reimagines process automation to deliver stable, scalable, and truly reliable results for enterprise workflows.
020
Salesforce AI Research @sfresearch.bsky.social · 07/04/2026
(1/5) 🎙️ Building Enterprise Realtime Voice Agents from Scratch: A Technical Tutorial Paper: bit.ly/4seq7Ee 25+ open-source speech-to-speech models exist, but none shows how to build a complete streaming voice agent with function calling.
120
Salesforce AI Research @sfresearch.bsky.social · 06/04/2026
1/5 Enterprise Sales Copilot: Enabling Real-Time AI Support with Automatic Information Retrieval in Live Sales Calls bit.ly/4tamShg 🎙️ Reps lose 25–65 sec per query searching CRM systems mid-call. SalesCopilot fixes that.
bit.ly
Enterprise Sales Copilot: Enabling Real-Time AI Support with Automatic Information Retrieval in Live Sales Calls
During live sales calls, customers frequently ask detailed product questions that require representatives to manually search internal databases and CRM systems. This process typically takes 25-65 seco...
130
Salesforce AI Research @sfresearch.bsky.social · 03/04/2026
Introducing GPA: GUI Process Automation Record one workflow demo. Replay it automatically — deterministic, fully local, and free. 100% success at ~10× the speed of Gemini 3 Pro's computer-use agent across 16 desktop tasks. 📝 Blog: salesforceairesearch.com/gpa 🔗 Paper: arxiv.org/abs/2604.01676
120
Salesforce AI Research @sfresearch.bsky.social · 01/04/2026
1/5 How Salesforce AI Research is Building Efficient RL Training for the Agentic Era Read full technical write-up: sforce.co/4bMqvUU
sforce.co
Building Efficient RL Training for the Agentic Era
Explore how Salesforce AI Research is optimizing Reinforcement Learning (RL) for the agentic era. Discover new frameworks for building efficient, scalable AI agents that power the future of autonomous...
120
Salesforce AI Research @sfresearch.bsky.social · 27/03/2026
Multimodal LLMs for Human-AI Interaction: Foundations, Agents, and Inclusive Applications mllm4haii.github.io Tutorial T4 at EACL 2026 covers how MLLMs extend beyond text to reason across visuals and GUIs. 📅 Sun, March 29 | 9:00-12:30 GMT+1 #FutureOfAI #EnterpriseAI #EACL2026
040
Salesforce AI Research @sfresearch.bsky.social · 27/03/2026
(1/5) VoiceAgentRAG: Solving the RAG Latency Bottleneck in Real-Time Voice Agents Using Dual-Agent Architectures Paper: bit.ly/4sMmqp4 🗣️ Voice agents need sub-200ms latency, but vector DB queries alone take 50–300ms.
120
Salesforce AI Research @sfresearch.bsky.social · 27/03/2026
"Many of the old rulebooks simply don't apply anymore." — Itai Asseo, VP of Salesforce AI Research, on how AI Foundry connects research to real business problems through rapid customer collaboration. sforce.co/4rTPhao #FutureOfAI #EnterpriseAI
sforce.co
Salesforce AI Research Launches AI Foundry to Accelerate System-Level Enterprise AI
As frontier models mature and become commodities, innovation in enterprise AI is happening at the system level, not the model level.
020
Salesforce AI Research @sfresearch.bsky.social · 26/03/2026
AI Foundry: Powering our big bets for research in 2026. sforce.co/4rTPhao #FutureOfAI #EnterpriseAI
030
Salesforce AI Research @sfresearch.bsky.social · 26/03/2026
"The problems that matter most for businesses don't live at the model level anymore. They live at the system level." — Silvio Savarese, EVP & Chief Scientist at @Salesforce sforce.co/4rTPhao #FutureOfAI #EnterpriseAI
sforce.co
Salesforce AI Research Launches AI Foundry to Accelerate System-Level Enterprise AI
As frontier models mature and become commodities, innovation in enterprise AI is happening at the system level, not the model level.
020
Salesforce AI Research @sfresearch.bsky.social · 26/03/2026
Introducing AI Foundry: Accelerating System-Level #EnterpriseAI sforce.co/4rTPhao As frontier models mature, the hardest enterprise challenges live at the system level. #AIFoundry brings together research, customers, and partners to develop and validate new AI capabilities. #FutureOfAI
sforce.co
Salesforce AI Research Launches AI Foundry to Accelerate System-Level Enterprise AI
As frontier models mature and become commodities, innovation in enterprise AI is happening at the system level, not the model level.
020
Salesforce AI Research @sfresearch.bsky.social · 24/03/2026
As AI moves toward multi-agent systems, the key challenge isn't just coordination — it's communication. Agents need their own semantic layer: shared protocols that let them negotiate and act on behalf of the organizations they represent.
040
Salesforce AI Research @sfresearch.bsky.social · 24/03/2026
Agents need more than natural language. They need a shared semantic layer to communicate and negotiate efficiently.
120
Salesforce AI Research @sfresearch.bsky.social · 24/03/2026
VIBEPASS: Can Vibe Coders Really Pass the Vibe Check? sforce.co/3NCuoT6 AI models score 90%+ on code generation, but under 50% on finding and fixing their own bugs. The gap between writing code and reasoning about faults remains wide. 🔍 #FutureOfAI
sforce.co
VIBEPASS: Can Vibe Coders Really Pass the Vibe Check?
VIBEPASS, a new benchmark, reveals a fundamental weakness in modern AI coding assistants: even with near-perfect scores on code generation tasks, frontier models falter when it comes to finding and fixing subtle bugs...
020
Salesforce AI Research @sfresearch.bsky.social · 24/03/2026
1/5 Agentic Confidence Calibration 📄 bit.ly/4kbjgs3 AI agents stay overconfident when they fail. HTC diagnoses the full execution trajectory, not just the final output.
bit.ly
Agentic Confidence Calibration
AI agents are rapidly advancing from passive language models to autonomous systems executing complex, multi-step tasks. Yet their overconfidence in failure remains a fundamental barrier to deployment ...
330
Salesforce AI Research @sfresearch.bsky.social · 23/03/2026
[1/12] 🚨 Your AI isn't as good at coding as you think. Frontier models are crushing benchmarks. Demos look magical. But here’s the uncomfortable truth: 👉 They’re great at writing code 👉 They’re bad at knowing when it’s wrong Enter: VIBEPASS 🧵
120