Moving from intent-based bots to proactive AI agents
OpenAI
Moving from intent-based bots to proactive AI agents.
Score: 37.4Confidence: 54%
View offerLoading the catalog…
THE AI OPPORTUNITY INDEX
Find your next AI tool. Explore free access, trials, and credits — all in one place.
OpenAI
Moving from intent-based bots to proactive AI agents.
Score: 37.4Confidence: 54%
View offerOpenAI
Can frontier LLMs earn $1 million from real-world freelance software engineering?
Score: 37.4Confidence: 54%
View offerOpenAI
We’ve implemented initial support for plugins in ChatGPT. Plugins are tools designed specifically for language models with safety as a core principle, and help ChatGPT access up-to-date information, run computations, or use third-party services.
Score: 37.4Confidence: 54%
View offerOpenAI
OpenAI banned accounts linked to a Russia-origin operation we dubbed “Stop News”, using AI to generate recidivist influence content targeting Africa and the UK.
Score: 37.4Confidence: 54%
View offerModal
We're excited to welcome Justin Dignelli to Modal. As VP of Sales, he will be leading our GTM efforts.
Score: 37.4Confidence: 54%
View offerOpenAI
Morgan Stanley uses AI evals to shape the future of financial services
Score: 37.4Confidence: 54%
View offerOpenAI
We’ve created GPT-4, the latest milestone in OpenAI’s effort in scaling up deep learning. GPT-4 is a large multimodal model (accepting image and text inputs, emitting text outputs) that, while less capable than humans in many real-world scenarios, exhibits human-level performance on various professional and academic benchmarks.
Score: 37.4Confidence: 54%
View offerOpenAI
This GPT-5 system card explains how a unified model routing system powers fast and smart responses using gpt-5-main, gpt-5-thinking, and lightweight versions like gpt-5-thinking-nano, optimized for different tasks and developer use.
Score: 37.4Confidence: 54%
View offerOpenAI
Developer registration for in-person attendance will open in the coming weeks and developers everywhere will be able to livestream the keynote.
Score: 37.4Confidence: 54%
View offerModal
Scale up smaller open models with search and evaluation to match frontier capabilities.
Score: 37.4Confidence: 54%
View offerOpenAI
We introduce MLE-bench, a benchmark for measuring how well AI agents perform at machine learning engineering.
Score: 37.4Confidence: 54%
View offerOpenAI
To support the safety of highly-capable AI systems, we are developing our approach to catastrophic risk preparedness, including building a Preparedness team and launching a challenge.
Score: 37.4Confidence: 54%
View offer