LiquidAI/ifstruct-v1.0
· benchmark:official, benchmark:eval-yaml, task_categories:text-generation
Home / 📊 Eval / Observability / Safety
14 View all · Daily curated AI / LLM open-source intelligence, with plain-language notes and license checks.
· benchmark:official, benchmark:eval-yaml, task_categories:text-generation
· task_categories:token-classification, language:ru, license:mit
· task_categories:text-to-image, task_categories:image-classification, task_categories:reinforcement-learning
· task_categories:text-to-speech, language:ja, license:mit
· gradio, leaderboard, asr
· region:us
Ultimate Claude Fable 5 Guide 2026: Use Cases, Integrations & Benchmarks
· benchmark:official, benchmark:eval-yaml, size_categories:n<1K
· benchmark:official, license:mit, size_categories:1K<n<10K
· language:en, size_categories:n<1K, format:parquet
· benchmark:official, benchmark:eval-yaml, task_categories:question-answering
The job search that runs on your machine. AI job application framework built on Claude Code: evaluate postings, tailor CVs, write cover letters, prep interviews. Fork it and own it.
· task_categories:text-generation, arxiv:2605.06754, region:us
Video-MME-v2: Towards the Next Stage in Benchmarks for Comprehensive Video Understanding