#001FlagEvalMM1 skills10724atualizado 2026-03-26100% do criadorskillocupaçãodescriçãoatualizadoflagevalmm-add-datasetDesenvolvedores de softwareIntegrate new evaluation datasets into FlagEvalMM as benchmark tasks. Use when adding a dataset from HuggingFace or other sources to FlagEvalMM, creating task configs, writing data processors, building custom evaluators, setting up prompt templates, or running evaluation benchmarks on VLMs. Trigger on: "add dataset to FlagEvalMM", "create a new task", "integrate benchmark", "evaluate model on [dataset]", "write process.py", "write evaluator", or any request involving the tasks/ directory of FlagEvalMM.2026-03-26