Collections
Discover the best community collections!
Collections trending this week
-
LEGENT: Open Platform for Embodied Agents
Paper • 2404.18243 • Published • 22 -
Revealing the Barriers of Language Agents in Planning
Paper • 2410.12409 • Published • 26 -
robbyant/lingbot-world-base-cam
Image-to-Video • Updated • 346 -
GameplayQA: A Benchmarking Framework for Decision-Dense POV-Synced Multi-Video Understanding of 3D Virtual Agents
Paper • 2603.24329 • Published • 25
-
HaninZ/DialogueEmotionClassification_DailyTalk
Viewer • Updated • 19k • 15 -
HaninZ/EnvironmentalSoundClassification_ESC50-HumanAndNonSpeechSounds_TTS
Viewer • Updated • 200 • 37 • 1 -
HaninZ/PronounciationEvaluationFluency_Speechocean762
Viewer • Updated • 2.5k • 54 • 1 -
HaninZ/MultiSpeakerDetection_LibriSpeech-TestClean_TTS
Viewer • Updated • 200 • 15
-
3D Room Layout Estimation LGT-Net
🏠160Generate 3D room layout from an RGB panorama image
-
Real-Time Latent Consistency Model Image-to-Image SD Turbo
🖼118Display a loading screen
-
IP-Adapter-FaceID
🧑1.02kGenerate AI images featuring your own face
-
CRM
📊306Generate 3D textured mesh from a single image
-
HaninZ/DialogueEmotionClassification_DailyTalk
Viewer • Updated • 19k • 15 -
HaninZ/EnvironmentalSoundClassification_ESC50-HumanAndNonSpeechSounds_TTS
Viewer • Updated • 200 • 37 • 1 -
HaninZ/PronounciationEvaluationFluency_Speechocean762
Viewer • Updated • 2.5k • 54 • 1 -
HaninZ/MultiSpeakerDetection_LibriSpeech-TestClean_TTS
Viewer • Updated • 200 • 15
-
3D Room Layout Estimation LGT-Net
🏠160Generate 3D room layout from an RGB panorama image
-
Real-Time Latent Consistency Model Image-to-Image SD Turbo
🖼118Display a loading screen
-
IP-Adapter-FaceID
🧑1.02kGenerate AI images featuring your own face
-
CRM
📊306Generate 3D textured mesh from a single image
-
LEGENT: Open Platform for Embodied Agents
Paper • 2404.18243 • Published • 22 -
Revealing the Barriers of Language Agents in Planning
Paper • 2410.12409 • Published • 26 -
robbyant/lingbot-world-base-cam
Image-to-Video • Updated • 346 -
GameplayQA: A Benchmarking Framework for Decision-Dense POV-Synced Multi-Video Understanding of 3D Virtual Agents
Paper • 2603.24329 • Published • 25