MiniMaxAI · Hugging Face 新模型· MiniMaxAI·· 2025-12-16AI 评分46
MiniMax VTP-Base-f16d64 视觉 tokenizer 模型发布
MiniMaxAI/VTP-Base-f16d64
AI 导读
MiniMax 发布 VTP 视觉 tokenizer 预训练框架,联合图文对比、自监督和重建损失,面向理解、重建与生成统一优化。VTP-L-f16d64 达到 78.2 zero-shot accuracy、85.7 linear probing、0.36 rFID,生成 FID-50K 为 2.81。
来源:MiniMaxAI · Hugging Face 新模型 · huggingface.co