이 페이지의 본문은 영어로 제공됩니다. 제목과 요약은 한국어로 번역되었습니다.
무엇인가
GPT-4o is a general-purpose model OpenAI released in May 2024; the “o” stands for omni, meaning text, image and audio are handled inside one model. Vision and speech understanding share a single forward pass, and spoken conversation runs at near-real-time latency. Where earlier systems stitched several models into a pipeline, GPT-4o trains multimodality end to end and simplifies the interaction entry point. It is offered through both the API and ChatGPT, and serves as OpenAI’s main multimodal general model.
기억할 만한 이유
It moved multimodality from “several models stitched into a pipeline” to “one model end to end” and cut spoken-dialogue latency to near human-conversation levels; general models have since treated native multimodality as the default target.
주요 사양
- Context window
- 128K tokens
- Modality
- Text, image, audio in; text, audio out
- Released
- 2024-05
- Open weights
- No
소속 능력
관련 개념
동종 제품
Claude
2023긴 문맥과 안전 정렬로 알려진 범용 대화 모델
Gemini
2023네이티브 멀티모달에 초장문 문맥을 다루는 범용 모델
Llama
2023개방형 가중치 노선을 주류로 만든 범용 모델 계열
Grok
2023소셜 플랫폼 데이터와 결합된 대화 모델
Mistral Large
2024유럽 오픈웨이트 연구소의 플래그십 상용 모델
Command R
2024검색 증강과 도구 호출을 위해 설계된 상용 모델
Kimi
2023긴 문맥 처리에 강한 중국어 대화 어시스턴트