29 repos
Python-based tools, frameworks, and evaluation systems for building and extending large language models and multimodal AI systems. The cluster spans model implementations (LLaVA, Otter), academic research infrastructure (gpt_academic), evaluation frameworks (evalplus, evoeval), and agent orchestration systems (NExT-GPT). Core themes include prompt engineering, model fine-tuning, multimodal vision-language capabilities, and systematic evaluation of LLM outputs.