EgoThink An evaluation benchmark for VLMs from the first-person perspective. Collection by SijieCheng Dec 6, 2024 1 Can Vision-Language Models Think from a First-Person Perspective? Paper • 2311.15596 • Published Nov 27, 2023 • 3 EgoThink/EgoThink Viewer • Updated Dec 6, 2023 • 700 • 205 • 6 VidEgoThink: Assessing Egocentric Video Understanding Capabilities for Embodied AI Paper • 2410.11623 • Published Oct 15, 2024 • 49
Can Vision-Language Models Think from a First-Person Perspective? Paper • 2311.15596 • Published Nov 27, 2023 • 3
VidEgoThink: Assessing Egocentric Video Understanding Capabilities for Embodied AI Paper • 2410.11623 • Published Oct 15, 2024 • 49
sdxl video Collection by angis Dec 5, 2023 - stabilityai/sdxl-turbo Text-to-Image • 3B • Updated Jul 10, 2024 • 1.04M • 2.61k
my projects Collection by lone-wolf-beta Dec 5, 2023 - Runtime error Question Answering Using Squad 💻
Code Collection by qwertyxasd Dec 5, 2023 - Magicoder: Source Code Is All You Need Paper • 2312.02120 • Published Dec 4, 2023 • 83
huggingface_llama Collection by AkiraL Dec 5, 2023 - meta-llama/Llama-2-70b-chat-hf Text Generation • 69B • Updated Apr 17, 2024 • 9.9k • 2.21k
Voiceover Collection by Robbin Dec 5, 2023 - Sleeping Agents 407 HierSpeech++ (Zero-shot TTS) ⚡ 407 Generate high-quality speech from text using a prompt audio
Sleeping Agents 407 HierSpeech++ (Zero-shot TTS) ⚡ 407 Generate high-quality speech from text using a prompt audio
paper Collection by janetwise Dec 5, 2023 - Magicoder: Source Code Is All You Need Paper • 2312.02120 • Published Dec 4, 2023 • 83 Mamba: Linear-Time Sequence Modeling with Selective State Spaces Paper • 2312.00752 • Published Dec 1, 2023 • 152
Mamba: Linear-Time Sequence Modeling with Selective State Spaces Paper • 2312.00752 • Published Dec 1, 2023 • 152
aiweb Collection by tomasdj Mar 4, 2025 - dropbox-dash/faster-whisper-large-v3-turbo Updated Nov 5, 2025 • 1.99M • 63
EgoThink An evaluation benchmark for VLMs from the first-person perspective. Collection by SijieCheng Dec 6, 2024 1 Can Vision-Language Models Think from a First-Person Perspective? Paper • 2311.15596 • Published Nov 27, 2023 • 3 EgoThink/EgoThink Viewer • Updated Dec 6, 2023 • 700 • 205 • 6 VidEgoThink: Assessing Egocentric Video Understanding Capabilities for Embodied AI Paper • 2410.11623 • Published Oct 15, 2024 • 49
Can Vision-Language Models Think from a First-Person Perspective? Paper • 2311.15596 • Published Nov 27, 2023 • 3
VidEgoThink: Assessing Egocentric Video Understanding Capabilities for Embodied AI Paper • 2410.11623 • Published Oct 15, 2024 • 49
Voiceover Collection by Robbin Dec 5, 2023 - Sleeping Agents 407 HierSpeech++ (Zero-shot TTS) ⚡ 407 Generate high-quality speech from text using a prompt audio
Sleeping Agents 407 HierSpeech++ (Zero-shot TTS) ⚡ 407 Generate high-quality speech from text using a prompt audio
sdxl video Collection by angis Dec 5, 2023 - stabilityai/sdxl-turbo Text-to-Image • 3B • Updated Jul 10, 2024 • 1.04M • 2.61k
paper Collection by janetwise Dec 5, 2023 - Magicoder: Source Code Is All You Need Paper • 2312.02120 • Published Dec 4, 2023 • 83 Mamba: Linear-Time Sequence Modeling with Selective State Spaces Paper • 2312.00752 • Published Dec 1, 2023 • 152
Mamba: Linear-Time Sequence Modeling with Selective State Spaces Paper • 2312.00752 • Published Dec 1, 2023 • 152
my projects Collection by lone-wolf-beta Dec 5, 2023 - Runtime error Question Answering Using Squad 💻
aiweb Collection by tomasdj Mar 4, 2025 - dropbox-dash/faster-whisper-large-v3-turbo Updated Nov 5, 2025 • 1.99M • 63
Code Collection by qwertyxasd Dec 5, 2023 - Magicoder: Source Code Is All You Need Paper • 2312.02120 • Published Dec 4, 2023 • 83
huggingface_llama Collection by AkiraL Dec 5, 2023 - meta-llama/Llama-2-70b-chat-hf Text Generation • 69B • Updated Apr 17, 2024 • 9.9k • 2.21k