Quick answer: Flan-T5 XL is Google's 3-billion-parameter instruction-tuned encoder-decoder model from the "Scaling Instruction-Finetuned Language Models" paper (Chung et al., 2022). It is open-weight (Apache 2.0) and is the largest Flan-T5 backbone commonly fine-tuned in agent and grounding research, including as the best-performing MindAct backbone in the Mind2Web benchmark.
Historical significance
Modern context
This model isn’t on any benchmark leaderboard yet.