AI & ML interests

None defined yet.

Recent Activity

prithivMLmodsย 
posted an update 5 days ago
view post
Post
3447
OneDecision-VisionGuard-Demo is now available on Hugging Face Spaces!

๐Ÿค— Space: prithivMLmods/OneDecision-VisionGuard-Demo

This demo showcases the OneDecision-VisionGuard family of multimodal image classification models for detecting NSFW and other sensitive visual content, with structured JSON reasoning, improved accuracy, and better handling of edge cases such as sensitive imagery, uncensored analysis, scene descriptions, and classification reasoning.

๐Ÿ“ฆ Models: 27B, 9B, 4B โ€” prithivMLmods/OneDecision-VisionGuard-27B-SFT, prithivMLmods/OneDecision-VisionGuard-9B-SFT, prithivMLmods/OneDecision-VisionGuard-4B-SFT

โ†—๏ธ Collection: https://huggingface.co/collections/prithivMLmods/onedecision-visionguard

To learn more, visit the app page or the respective model pages.
arudradeyย 
posted an update 15 days ago
prithivMLmodsย 
posted an update 18 days ago
view post
Post
3914
Qwen-Image-2.1 Plug and Play LoRA App is now live on Hugging Face Spaces.

๐Ÿ”— Space: prithivMLmods/Qwen-Image-2.1-LoRAs-PnP

It supports standard inference, 4-step Turbo inference, custom LoRA lazy repacks, and LoRA Plug and Play (PnP), all in one setting!

๐Ÿ”— Qwen-Image-2.1 Image-to-Image LoRAs: https://huggingface.co/collections/prithivMLmods/qwen-image-21-image-to-image-loras

๐Ÿ”— GitHub: https://github.com/PRITHIVSAKTHIUR/Qwen-Image-2.1-LoRAs-PnP

To learn more, visit the app page or the respective model pages.
NILKNARFGonzoย 
posted an update 19 days ago
view post
Post
107
just recieved my stack of 10 floppy disks - you know what that means

floppyx4 is canceled, floppyx10 is next

here's the intended specs:
- official tokenizer (the actual tokenizer for gpt-2)
- actual gpu training (barely)
- sharegpt (if i can afford it computationally)
- full thing fitting on 10 floppy disks (not just the safetensors file)
- and if needed different arch (like llama)

also unsloth on a gpu from 2015 is insane
  • 3 replies
ยท
NILKNARFGonzoย 
posted an update 21 days ago
view post
Post
85
i think someone posted my password and ip on some platform and im being hacked left and right
  • 5 replies
ยท
NILKNARFGonzoย 
posted an update 22 days ago
view post
Post
3775
get played unsloth

gemma just deleted its own model runner with DeepSeek Harness

shoutout to deepseek and unsloth
  • 11 replies
ยท
prithivMLmodsย 
posted an update 26 days ago
view post
Post
845
VisionGuardrail EVO-2, a multimodal image-classification content-safety model based on Qwen/Qwen3.8-27B, is now available on the Hub!

Stricter image classification than before, with a dense 27-billion-parameter multimodal model, more precise reasoning, and improved captions for classifying visual media.

โž  Models: prithivMLmods/VisionGuardrail-Evo2-27B, prithivMLmods/VisionGuardrail-Evo2-27B-GGUF

โž  Collection: https://huggingface.co/collections/prithivMLmods/visionguardrail-evo2

โž  Previous Models: https://huggingface.co/collections/prithivMLmods/visionguardrail-collection

โคท To learn more, visit the app page or the respective model pages.
prithivMLmodsย 
posted an update 29 days ago
view post
Post
492
Scribble-Board-Fast is a sketch-to-image workspace powered by Klein-9B, transforming doodles, brush strokes, stickers, and uploaded images into high-fidelity visuals with 4-step distilled sampling.

> Space: prithivMLmods/Scribble-Board-Fast
> GitHub: https://github.com/PRITHIVSAKTHIUR/Scribble-Board-Fast

> To learn more, visit the app page or the respective model pages.
NILKNARFGonzoย 
posted an update 30 days ago
view post
Post
120
Open-source is not going away anytime soon.

There's a handful of open models like Qwen Image, Flux, Wan, and others on AI image generators like VisualGPT, Pixlr, Free AI, and more. And here's the catch - they're free.

Hugging Face inference costs money just to generate simple images. Things like VisualGPT still have your favorite models for free.

GPT Image isn't worth it - and neither is HF inference. The real way to use open models is the things you closed-source third-party lovers already use.
  • 1 reply
ยท
NILKNARFGonzoย 
posted an update about 1 month ago
view post
Post
95
guys! if ur part of a team org i think you can post

so yeah

@GGUFGuy i figured out why u can post (bc of HuggingScience)
  • 1 reply
ยท
NILKNARFGonzoย 
in open-acc/README about 1 month ago

help

#13 opened about 1 month ago by
NILKNARFGonzo
prithivMLmodsย 
posted an update about 1 month ago
view post
Post
3872
VisionGuardrail, a multimodal content-safety classifier based on Qwen3.5, is now available on Hugging Face in 4B and 9B variants. It is a direct upgrade to ImageShield-MMCF, providing improved parental controls through conservative visual content-safety filtering.

More About:
โž  hf.co/blog โ€” https://huggingface.co/blog/prithivMLmods/vision-guardrail-mini-blog

โž  Models:
โœฆ VisionGuardrail-4B: prithivMLmods/VisionGuardrail-4B
โœฆ VisionGuardrail-9B: prithivMLmods/VisionGuardrail-9B

โž  Dataset:
โœฆ ImageShield-Guardrail-Pro: prithivMLmods/ImageShield-Guardrail-Pro

โคท To learn more, visit the app page or the respective model pages.
prithivMLmodsย 
posted an update about 2 months ago
view post
Post
3109
ImageShield-MMCF โ€” Multimodal Content Filter is a multimodal content-safety classifier built on top of Qwen3.5 and is now available on Hugging Face!

This is the preview initial version (v1.0) of the model, designed to classify visual content as Safe or Unsafe, with a particular focus on detecting Not Safe for Work (NSFW) and other potentially sensitive visual content.

The demo is implemented in the prithivMLmods/opencaption-4b-vl-sft Space, which serves as an active content-safety layer for computer vision tasks. It helps block Not Safe for Work (NSFW) content generation and paves the way for more meaningful and responsible creativity.

โŠน ImageShield-MMCF-0.8B: prithivMLmods/ImageShield-MMCF-0.8B
โŠน ImageShield-MMCF-2B: prithivMLmods/ImageShield-MMCF-2B
  • 2 replies
ยท
prithivMLmodsย 
posted an update about 2 months ago
view post
Post
5295
The Qwen3.8 27B demo for object grounding is now available on Hugging Face Spaces.

It features three tasks: Object Detection (Bounding Boxes), Point Localization (Keypoints), and Spatial Guidance (Path Mapping).

Try it now: prithivMLmods/Qwen3.8-27B-Object-Detection
Nymboย 
posted an update 2 months ago
view post
Post
2443
Anthropic gave me six months of Claude Max 20x through the Claude for Open Source program, granted based on my Hugging Face work. Thank you
Anthropic
for supporting open source.

So far I've been pointing it at Markdown Minimap, an Obsidian plugin that adds a scrollable IDE-style minimap to your notes. This week I've been clearing a backlog of user-reported issues on it, with Claude often handling them end to end.

https://github.com/Nymbo/Markdown-Minimap โ€” issues and PRs welcome.
prithivMLmodsย 
posted an update 2 months ago
view post
Post
5575
Made a demo for Text/Image-to-3D Video and Image-to-3D Video asset generation using TRELLIS.2. It is paired with Z-Image-Turbo to accelerate the input image preprocessing pipeline, streamlining the Image-to-3D workflow. The generated GLB (GL Transmission Format) files are converted into MP4 (MPEG-4) videos, making them easy to preview and share. Try it now on Hugging Face Spaces.๐Ÿค—

โž  Image-to-3D-Video-Asset-Generator: prithivMLmods/Image-to-3D-Video-Asset-Generator
โž  collection: https://huggingface.co/collections/prithivMLmods/multimodal-implementations
โž  github: https://github.com/PRITHIVSAKTHIUR/Image-to-3D-Video-Asset-Generator

โคท To learn more, visit the app page or the respective model pages.
julien-cย 
posted an update 3 months ago
view post
Post
6029
who's working on an NVFP4 version of Kimi-K3?
  • 4 replies
ยท
Nymboย 
posted an update 3 months ago
view post
Post
6133
Introducing Inflect-v2, two exceptionally small, open-weight English TTS models at just 3.9M and 9.3M parameters. Both generate speech multiple times faster than real-time on CPU. Despite their size, Inflect-v2 delivers quality that is competitive with much larger lightweight TTS systems, including KittenTTS, Piper, and Supertonic-3.

CPU, CUDA, PyTorch, and ONNX are supported. Apache 2.0.

See it for yourselves:
owensong/Inflect-Micro-v2
owensong/Inflect-Nano-v2

Try the Demos:
Nymbo/Inflect-TTS (unlimited CPU usage)
owensong/Inflect-v2 (ultra-fast ZeroGPU usage)
  • 6 replies
ยท