Average Ratings 0 Ratings
Average Ratings 1 Rating
Description
Claude Sonnet 4.5 represents Anthropic's latest advancement in AI, crafted to thrive in extended coding environments, complex workflows, and heavy computational tasks while prioritizing safety and alignment. It sets new benchmarks with its top-tier performance on the SWE-bench Verified benchmark for software engineering and excels in the OSWorld benchmark for computer usage, demonstrating an impressive capacity to maintain concentration for over 30 hours on intricate, multi-step assignments. Enhancements in tool management, memory capabilities, and context interpretation empower the model to engage in more advanced reasoning, leading to a better grasp of various fields, including finance, law, and STEM, as well as a deeper understanding of coding intricacies. The system incorporates features for context editing and memory management, facilitating prolonged dialogues or multi-agent collaborations, while it also permits code execution and the generation of files within Claude applications. Deployed at AI Safety Level 3 (ASL-3), Sonnet 4.5 is equipped with classifiers that guard against inputs or outputs related to hazardous domains and includes defenses against prompt injection, ensuring a more secure interaction. This model signifies a significant leap forward in the intelligent automation of complex tasks, aiming to reshape how users engage with AI technologies.
Description
GPT-6.1 Sol is an upgraded OpenAI model that combines advanced intelligence with lower operating costs for coding, professional knowledge work, computer use, scientific research, and autonomous agents. OpenAI positions it as offering near-GPT-6 Astra intelligence at one-fifth of Astra's standard input and output token prices. The model delivers substantial improvements over GPT-6 Sol in software engineering, complex document understanding, business automation, and long-horizon computer-use workflows. On DeepSWE v1.1, GPT-6.1 Sol matches GPT-6 Astra at approximately one-fifth of the cost and surpasses GPT-6 Sol's highest score by 6.4 percentage points at lower reasoning effort. On AutomationBench, it scores 4.8 percentage points higher than GPT-6 Sol at the same reasoning setting and 2.2 points above Opus 5.5 at medium reasoning effort. GPT-6.1 Sol also improves computer use, coming within 2.1 percentage points of GPT-6 Astra on the OSWorld 2.0 offline set at maximum reasoning effort while costing roughly one-seventh as much per task. For scientific workflows, the model more than doubles GPT-6 Sol's Terminal-Bench Science 0.1 score at maximum effort while reducing average task cost by more than half. Factuality has also improved, with the share of responses containing a factual error at low reasoning effort falling from 11.4% with GPT-6 Sol to 7.7% with GPT-6.1 Sol on OpenAI's difficult error-focused evaluation. Developers can access GPT-6.1 Sol through the OpenAI API for $2 per million input tokens, $0.10 per million cached input tokens, and $10 per million output tokens, while eligible users can access it through ChatGPT Work and Codex.
API Access
Has API
Yes
API Access
Has API
Yes
Integrations
Amazon Bedrock
Yes
Augment Code
Yes
CSS
Yes
Charlie
Yes
Dessix
Yes
Devin
Yes
Doraverse
Yes
GitHub
Yes
GitHub Copilot
Yes
HTML
Yes
Integrations
Amazon Bedrock
Yes
Augment Code
Yes
CSS
Yes
Charlie
Yes
Dessix
Yes
Devin
Yes
Doraverse
Yes
GitHub
Yes
GitHub Copilot
Yes
HTML
Yes
Pricing Details
No price information available.
Free Trial
No
Free Version
No
Pricing Details
$2 per 1M tokens (input)
Input: $2 per 1 million tokens
Output: $10 per 1 million tokens
Cached Input: $0.10 per 1 million cached input tokens
Output: $10 per 1 million tokens
Cached Input: $0.10 per 1 million cached input tokens
Free Trial
No
Free Version
No
Deployment
Web-Based
Yes
On-Premises
No
iPhone App
Yes
iPad App
No
Android App
Yes
Windows
No
Mac
No
Linux
No
Chromebook
No
Deployment
Web-Based
Yes
On-Premises
No
iPhone App
No
iPad App
No
Android App
No
Windows
No
Mac
No
Linux
No
Chromebook
No
Customer Support
Business Hours
Yes
Live Rep (24/7)
No
Online Support
Yes
Customer Support
Business Hours
No
Live Rep (24/7)
No
Online Support
Yes
Types of Training
Training Docs
Yes
Webinars
Yes
Live Training (Online)
No
In Person
Yes
Types of Training
Training Docs
Yes
Webinars
No
Live Training (Online)
No
In Person
No
Vendor Details
Company Name
Anthropic
Founded
2021
Country
United States
Website
claude.ai
Vendor Details
Company Name
OpenAI
Founded
2015
Country
United States
Website
openai.com