Skip to main content

A library of foundation models in computer vision and multi-modal learning.

5
GitHub Stars
94
Curated Resources
9
Categories
17 hours ago
Last Refreshed
PretrainingGenerationUnified Architecture for VisionInstruction TuningRLHFChat ModelsVisual Chat ModelsDatasetsEvaluation

Use this list with your AI agent

Add the Context Awesome MCP server to Claude, Cursor, or any MCP client, then ask:

"Show me chinese support resources from awesome-foundation-models"

Installation instructions →

What's inside

Chat Models

Evaluation

Resources

Unified Architecture for Vision

Showing a sample of 94 resources. View the full list on GitHub →