Scaling In-Context Reinforcement Learning, ChartMimic's AI Benchmark, Multimodal Document Comprehension, Long Context Reasoning Challenges
MP3•Laman utama episod
Manage episode 424766685 series 3568650
Kandungan disediakan oleh PocketPod. Semua kandungan podcast termasuk episod, grafik dan perihalan podcast dimuat naik dan disediakan terus oleh PocketPod atau rakan kongsi platform podcast mereka. Jika anda percaya seseorang menggunakan karya berhak cipta anda tanpa kebenaran anda, anda boleh mengikuti proses yang digariskan di sini https://ms.player.fm/legal.
XLand-100B: A Large-Scale Multi-Task Dataset for In-Context Reinforcement Learning Make It Count: Text-to-Image Generation with an Accurate Number of Objects ChartMimic: Evaluating LMM's Cross-Modal Reasoning Capability via Chart-to-Code Generation Needle In A Multimodal Haystack BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack
…
continue reading
70 episod