llms Collection by Wolverine001 May 21, 2024 - meta-llama/Llama-2-7b-hf Text Generation • 7B • Updated Apr 17, 2024 • 794k • 2.41k CompVis/ldm-celebahq-256 Unconditional Image Generation • Updated Jul 28, 2022 • 1.31k • 51
HGRN2 HGRN2: Gated Linear RNNs with State Expansion Collection by OpenNLPLab123 Jun 25, 2024 2 HGRN2: Gated Linear RNNs with State Expansion Paper • 2404.07904 • Published Apr 11, 2024 • 21 Scaling Laws for Linear Complexity Language Models Paper • 2406.16690 • Published Jun 24, 2024 • 23
HGRN Hierarchically Gated Recurrent Neural Network for Sequence Modeling Collection by OpenNLPLab123 Apr 11, 2024 - OpenNLPLab/HGRN-150M Text Generation • Updated Nov 10, 2023 • 36 • 2 OpenNLPLab/HGRN-355M Text Generation • Updated Nov 10, 2023 • 21 • 2 OpenNLPLab/HGRN-1B Text Generation • Updated Nov 10, 2023 • 21 • 8 Hierarchically Gated Recurrent Neural Network for Sequence Modeling Paper • 2311.04823 • Published Nov 8, 2023 • 2
Hierarchically Gated Recurrent Neural Network for Sequence Modeling Paper • 2311.04823 • Published Nov 8, 2023 • 2
instruction tuning Collection by mok0102 Apr 11, 2024 - NEFTune: Noisy Embeddings Improve Instruction Finetuning Paper • 2310.05914 • Published Oct 9, 2023 • 14
NEFTune: Noisy Embeddings Improve Instruction Finetuning Paper • 2310.05914 • Published Oct 9, 2023 • 14
KPU Benchmarks v1 Benchmarks used to test the KPU for the Tech Report Collection by MAISAAI Oct 15, 2024 -
LLM Collection by erictsai Apr 16, 2024 - lmsys/vicuna-7b-v1.5 Text Generation • Updated Mar 13, 2024 • 36.4k • 403 taide/TAIDE-LX-7B-Chat Text Generation • 7B • Updated May 21, 2024 • 904 • 150
paper Collection by kimv2v2 Apr 29, 2024 - DreamScene360: Unconstrained Text-to-3D Scene Generation with Panoramic Gaussian Splatting Paper • 2404.06903 • Published Apr 10, 2024 • 21 PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning Paper • 2404.16994 • Published Apr 25, 2024 • 39
DreamScene360: Unconstrained Text-to-3D Scene Generation with Panoramic Gaussian Splatting Paper • 2404.06903 • Published Apr 10, 2024 • 21
PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning Paper • 2404.16994 • Published Apr 25, 2024 • 39
llms Collection by Wolverine001 May 21, 2024 - meta-llama/Llama-2-7b-hf Text Generation • 7B • Updated Apr 17, 2024 • 794k • 2.41k CompVis/ldm-celebahq-256 Unconditional Image Generation • Updated Jul 28, 2022 • 1.31k • 51
HGRN2 HGRN2: Gated Linear RNNs with State Expansion Collection by OpenNLPLab123 Jun 25, 2024 2 HGRN2: Gated Linear RNNs with State Expansion Paper • 2404.07904 • Published Apr 11, 2024 • 21 Scaling Laws for Linear Complexity Language Models Paper • 2406.16690 • Published Jun 24, 2024 • 23
KPU Benchmarks v1 Benchmarks used to test the KPU for the Tech Report Collection by MAISAAI Oct 15, 2024 -
HGRN Hierarchically Gated Recurrent Neural Network for Sequence Modeling Collection by OpenNLPLab123 Apr 11, 2024 - OpenNLPLab/HGRN-150M Text Generation • Updated Nov 10, 2023 • 36 • 2 OpenNLPLab/HGRN-355M Text Generation • Updated Nov 10, 2023 • 21 • 2 OpenNLPLab/HGRN-1B Text Generation • Updated Nov 10, 2023 • 21 • 8 Hierarchically Gated Recurrent Neural Network for Sequence Modeling Paper • 2311.04823 • Published Nov 8, 2023 • 2
Hierarchically Gated Recurrent Neural Network for Sequence Modeling Paper • 2311.04823 • Published Nov 8, 2023 • 2
LLM Collection by erictsai Apr 16, 2024 - lmsys/vicuna-7b-v1.5 Text Generation • Updated Mar 13, 2024 • 36.4k • 403 taide/TAIDE-LX-7B-Chat Text Generation • 7B • Updated May 21, 2024 • 904 • 150
instruction tuning Collection by mok0102 Apr 11, 2024 - NEFTune: Noisy Embeddings Improve Instruction Finetuning Paper • 2310.05914 • Published Oct 9, 2023 • 14
NEFTune: Noisy Embeddings Improve Instruction Finetuning Paper • 2310.05914 • Published Oct 9, 2023 • 14
paper Collection by kimv2v2 Apr 29, 2024 - DreamScene360: Unconstrained Text-to-3D Scene Generation with Panoramic Gaussian Splatting Paper • 2404.06903 • Published Apr 10, 2024 • 21 PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning Paper • 2404.16994 • Published Apr 25, 2024 • 39
DreamScene360: Unconstrained Text-to-3D Scene Generation with Panoramic Gaussian Splatting Paper • 2404.06903 • Published Apr 10, 2024 • 21
PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning Paper • 2404.16994 • Published Apr 25, 2024 • 39