AI Research Atlas
Technical report

Multimodal models

GPT-4 reports broader multimodal capability

A large model accepted image and text inputs and generated text outputs.

OpenAI

AI topics

Explore related entries. Larger tags appear on more entries.

The contribution

The report describes GPT-4 evaluation across a range of professional, academic, and language benchmarks, alongside limitations and safety work. It documents a multimodal model while withholding many details of its architecture and training.

What this does not establish

Limited technical disclosure constrains independent reproduction. Strong benchmark results do not establish general intelligence or consistent factual accuracy.

Why this date?

The technical report was first submitted on 15 March 2023. GPT-4o is a separate, later model.

This entry follows the linked publication. Read the source and date conventions.

Comments

Discuss this research, ask a question, or suggest a correction. Comments appear after the site owner approves them.

Loading comments…

Sign in with ChatGPT to comment

Use your OpenAI account. Published comments show the display name you choose, not your account email.