Key Specifications

Vendoranthropic
Version3.5-sonnet
Release Date2024-06-20
Context Window200000 tokens
Input Modalitiestext, image
Output Modalitiestext
LicenseProprietary
Documentationhttps://docs.anthropic.com/claude/docs

Benchmark Performance

BenchmarkScoreUnitEvaluated AtNotesSource
MMLU88.7%2024-06-205-shotview
HUMANEVAL92pass@12024-06-20view
GSM8K96.4%2024-06-200-shot CoTview
MATH71.1%2024-06-200-shot CoTview
BBH84.5%2024-06-203-shot CoTview

Pricing

TierPriceCurrency
Input$3 / MtokUSD
Output$15 / MtokUSD
Cache Read$0.3 / MtokUSD
Cache Write$3.75 / MtokUSD

Source: https://www.anthropic.com/pricing · as of 2024-08-01

Compliance

  • Data Residency: US
  • SOC2: ✓
  • HIPAA: ✗
  • GDPR: ✓
  • ISO 27001: ✗

Claude 3.5 Sonnet

모델 개요

Anthropic Claude 3.5 Sonnet 模型,200K 上下文窗口,在编码、视觉推理与长文本理解方面表现突出,平衡了智能与速度。

핵심 사양

공급업체버전출시일컨텍스트 창입력 모달리티출력 모달리티라이선스
Anthropic3.5-sonnet2024-06-20200Ktext, imagetextProprietary

벤치마크 성능

벤치마크점수단위비고
MMLU (Massive Multitask Language Understanding)88.7%5-shot
HumanEval92.0pass@1
GSM8K (Grade School Math 8K)96.4%0-shot CoT
MATH71.1%0-shot CoT
BBH (BIG-Bench Hard)84.5%3-shot CoT

가격

입력출력캐시 읽기캐시 쓰기

백만 토큰당

강점

  • MMLU score 88.7, strong knowledge reasoning.
  • HumanEval 92.0, excellent code generation.
  • GSM8K 96.4, robust math reasoning.
  • 上下文窗口 200K,支持长文本。

약점

  • 闭源专有模型,不支持自托管。

사용 사례

  • 代码生成与调试
  • 长文档摘要
  • Agent 工作流与工具调用

참고문헌