AI Research Atlas
Technical report

Efficient models

Mistral 7B focuses on efficient attention

Grouped-query and sliding-window attention supported an efficient language-model architecture.

Albert Q. Jiang and colleagues

AI topics

Explore related entries. Larger tags appear on more entries.

The contribution

The report presents Mistral 7B and evaluates it against other models on a range of tasks. Its architecture uses grouped-query attention and sliding-window attention to improve inference efficiency and manage attention computation.

What this does not establish

Benchmark comparisons describe the models and evaluation setups in the report. They should not be read as a permanent ranking of available systems.

Why this date?

The paper was submitted on 10 October 2023; the model announcement was on 27 September 2023.

This entry follows the linked publication. Read the source and date conventions.

Comments

Discuss this research, ask a question, or suggest a correction. Comments appear after the site owner approves them.

Loading comments…

Sign in with ChatGPT to comment

Use your OpenAI account. Published comments show the display name you choose, not your account email.