Marco
AI & ML interests
Recent Activity
Organizations
-
stepfun-ai/GOT-OCR-2.0-hf
Image-Text-to-Text • 0.6B • Updated • 149k • 237 - Running on ZeroAgents84
GOT OCR Transformers
📷84Demo of GOT-OCR 2.0's Transformers implementation
-
allenai/olmOCR-7B-0225-preview
Image-Text-to-Text • 8B • Updated • 4.08k • 707 -
allenai/olmOCR-mix-0225
Viewer • Updated • 259k • 377 • 171
- Running559
DeepSeek-R1 WebGPU
🧠559Next-generation reasoning model that runs locally in-browser
- PausedAgents101
Qwen2.5-1M Demo
💻101Ask questions about your uploaded documents
-
mistralai/Mistral-Small-24B-Base-2501
24B • Updated • 3.65k • 264 -
deepseek-ai/deepseek-vl2-small
Image-Text-to-Text • 16B • Updated • 5.66k • 180
- RunningAgentsFeatured270
Qwen3 Omni Demo
⚡270Chat with AI using text, audio, images, or video
- RunningAgents66
Qwen3 Omni Captioner Demo
🐠66Generate a caption for any uploaded or recorded audio
-
Qwen/Qwen3-Omni-30B-A3B-Thinking
Any-to-Any • 32B • Updated • 444k • 313 -
Qwen/Qwen3-Omni-30B-A3B-Instruct
Any-to-Any • 35B • Updated • 1.48M • 965
- RunningMCP135
Consilium MCP Server
🏢135Multi-AI Expert Consensus Platform
- Runtime errorMCP2
MCP Hackathon Deepfake Watchdog
🛡2Upload your image and/or voice to scan for deepfake misuse o
- Runtime error36
VulnBuster
🛡36AI Security Agent: Multi-MCP Code Vulnerability Scanner
- RunningMCP200
AI Marketing Content Generator
🎨200An AI-powered tool made for content creators and marketers
-
nvidia/parakeet-tdt-0.6b-v2
Automatic Speech Recognition • Updated • 736k • 1.53k - Running on ZeroAgentsFeatured479
Parakeet-TDT-0.6b-V2
479Transcribe audio files with timestamps and downloadable subtitles
- Runtime errorAgents33
Blazing Fast Whisper
👁33Blazing Fast Whisper Deployed on HF Inference Endpoints
- Running on CPU UpgradeAgentsFeatured1.42k
Open ASR Leaderboard
🏆1.42kExplore speech model benchmarks across languages and datasets
- Running on T4Agents147
RF-DETR
🔥147SOTA real-time object detection model
- Running on CPU UpgradeAgents50
YOLO ARENA
🏟50compare performance of top object detectors
- Running on ZeroAgentsFeatured93
D-Fine - SOTA Real-Time Object Detector
⚡93Object Detection on Images and Video
- Running on ZeroMCP31
Gaze LLE
👀31Gaze Target Estimation
-
flax-community/t5-recipe-generation
Text Generation • 0.2B • Updated • 319 • 77 -
numind/NuExtract-1.5
Text Generation • 4B • Updated • 3.94k • 247 - RunningAgents11
Signature Detection
👁11Handwritten Signature Detection
-
manycore-research/SpatialLM-Llama-1B
Text Generation • 1B • Updated • 841 • 993
- Running on ZeroMCPFeatured606
LatentSync
👄606Audio Conditioned LipSync with Latent Diffusion Models
- PausedAgents228
BEN2
🚀228Remove background from images and videos
- Running on ZeroAgents81
SmolVLM
📊81Ask questions about images or videos and get answers
- Build errorAgents59
SmolVLM2 HighlightGenerator
🐨59Generate video highlights from uploaded video
-
NexaAI/Qwen2-Audio-7B-GGUF
Audio-Text-to-Text • 8B • Updated • 1.69k • 170 -
kyutai/hibiki-2b-pytorch-bf16
Translation • Updated • 454 • 63 -
Zyphra/Zonos-v0.1-hybrid
Text-to-Speech • 2B • Updated • 1.45k • 1.11k - Running on ZeroAgentsFeatured688
Di♪♪Rhythm
🎶688Blazingly Fast and Embarrassingly Simple Song Generation
-
onnx-community/Kokoro-82M-ONNX
Text-to-Speech • Updated • 51.6k • 179 - Running223
Kokoro Text-to-Speech
🗣223High-quality speech synthesis powered by Kokoro TTS
-
NexaAI/Qwen2-Audio-7B-GGUF
Audio-Text-to-Text • 8B • Updated • 1.69k • 170 -
jonatasgrosman/wav2vec2-large-xlsr-53-english
Automatic Speech Recognition • 0.3B • Updated • 55.3k • 477
- RunningAgentsFeatured270
Qwen3 Omni Demo
⚡270Chat with AI using text, audio, images, or video
- RunningAgents66
Qwen3 Omni Captioner Demo
🐠66Generate a caption for any uploaded or recorded audio
-
Qwen/Qwen3-Omni-30B-A3B-Thinking
Any-to-Any • 32B • Updated • 444k • 313 -
Qwen/Qwen3-Omni-30B-A3B-Instruct
Any-to-Any • 35B • Updated • 1.48M • 965
- RunningMCP135
Consilium MCP Server
🏢135Multi-AI Expert Consensus Platform
- Runtime errorMCP2
MCP Hackathon Deepfake Watchdog
🛡2Upload your image and/or voice to scan for deepfake misuse o
- Runtime error36
VulnBuster
🛡36AI Security Agent: Multi-MCP Code Vulnerability Scanner
- RunningMCP200
AI Marketing Content Generator
🎨200An AI-powered tool made for content creators and marketers
-
nvidia/parakeet-tdt-0.6b-v2
Automatic Speech Recognition • Updated • 736k • 1.53k - Running on ZeroAgentsFeatured479
Parakeet-TDT-0.6b-V2
479Transcribe audio files with timestamps and downloadable subtitles
- Runtime errorAgents33
Blazing Fast Whisper
👁33Blazing Fast Whisper Deployed on HF Inference Endpoints
- Running on CPU UpgradeAgentsFeatured1.42k
Open ASR Leaderboard
🏆1.42kExplore speech model benchmarks across languages and datasets
- Running on T4Agents147
RF-DETR
🔥147SOTA real-time object detection model
- Running on CPU UpgradeAgents50
YOLO ARENA
🏟50compare performance of top object detectors
- Running on ZeroAgentsFeatured93
D-Fine - SOTA Real-Time Object Detector
⚡93Object Detection on Images and Video
- Running on ZeroMCP31
Gaze LLE
👀31Gaze Target Estimation
-
flax-community/t5-recipe-generation
Text Generation • 0.2B • Updated • 319 • 77 -
numind/NuExtract-1.5
Text Generation • 4B • Updated • 3.94k • 247 - RunningAgents11
Signature Detection
👁11Handwritten Signature Detection
-
manycore-research/SpatialLM-Llama-1B
Text Generation • 1B • Updated • 841 • 993
-
stepfun-ai/GOT-OCR-2.0-hf
Image-Text-to-Text • 0.6B • Updated • 149k • 237 - Running on ZeroAgents84
GOT OCR Transformers
📷84Demo of GOT-OCR 2.0's Transformers implementation
-
allenai/olmOCR-7B-0225-preview
Image-Text-to-Text • 8B • Updated • 4.08k • 707 -
allenai/olmOCR-mix-0225
Viewer • Updated • 259k • 377 • 171
- Running on ZeroMCPFeatured606
LatentSync
👄606Audio Conditioned LipSync with Latent Diffusion Models
- PausedAgents228
BEN2
🚀228Remove background from images and videos
- Running on ZeroAgents81
SmolVLM
📊81Ask questions about images or videos and get answers
- Build errorAgents59
SmolVLM2 HighlightGenerator
🐨59Generate video highlights from uploaded video
-
NexaAI/Qwen2-Audio-7B-GGUF
Audio-Text-to-Text • 8B • Updated • 1.69k • 170 -
kyutai/hibiki-2b-pytorch-bf16
Translation • Updated • 454 • 63 -
Zyphra/Zonos-v0.1-hybrid
Text-to-Speech • 2B • Updated • 1.45k • 1.11k - Running on ZeroAgentsFeatured688
Di♪♪Rhythm
🎶688Blazingly Fast and Embarrassingly Simple Song Generation
- Running559
DeepSeek-R1 WebGPU
🧠559Next-generation reasoning model that runs locally in-browser
- PausedAgents101
Qwen2.5-1M Demo
💻101Ask questions about your uploaded documents
-
mistralai/Mistral-Small-24B-Base-2501
24B • Updated • 3.65k • 264 -
deepseek-ai/deepseek-vl2-small
Image-Text-to-Text • 16B • Updated • 5.66k • 180
-
onnx-community/Kokoro-82M-ONNX
Text-to-Speech • Updated • 51.6k • 179 - Running223
Kokoro Text-to-Speech
🗣223High-quality speech synthesis powered by Kokoro TTS
-
NexaAI/Qwen2-Audio-7B-GGUF
Audio-Text-to-Text • 8B • Updated • 1.69k • 170 -
jonatasgrosman/wav2vec2-large-xlsr-53-english
Automatic Speech Recognition • 0.3B • Updated • 55.3k • 477