Daniel Rosehill PRO
AI & ML interests
Recent Activity
Organizations
-
imvladikon/wav2vec2-large-xlsr-53-hebrew
Automatic Speech Recognition • 0.3B • Updated • 458 • 7 -
Mizurodp/wav2vec2-large-xls-r-300m-hebrew-colab
Automatic Speech Recognition • Updated • 68 • 1 -
imvladikon/wav2vec2-xls-r-300m-lm-hebrew
Automatic Speech Recognition • 0.3B • Updated • 17 • 4 -
imvladikon/wav2vec2-xls-r-1b-hebrew
Automatic Speech Recognition • 1.0B • Updated • 9 • 2
-
danielrosehill/daniel_whisper_finetune_large_v3_turbo_v2
Automatic Speech Recognition • 0.8B • Updated • 1 -
danielrosehill/daniel_whisper_finetune_medium_v2
Automatic Speech Recognition • 0.8B • Updated • 2 -
danielrosehill/daniel_whisper_finetune_tiny_v2
Automatic Speech Recognition • 37.8M • Updated • 1 -
danielrosehill/daniel_whisper_finetune_base_v2
Automatic Speech Recognition • 72.6M • Updated • 7
-
nvidia/parakeet-tdt-0.6b-v3
Automatic Speech Recognition • 0.6B • Updated • 402k • 803 -
ibm-granite/granite-speech-3.3-8b
Automatic Speech Recognition • 9B • Updated • 81.8k • 166 -
facebook/seamless-m4t-v2-large
Automatic Speech Recognition • 2B • Updated • 74.9k • 972 -
facebook/wav2vec2-base-960h
Automatic Speech Recognition • 94.4M • Updated • 1.2M • 396
-
futo-org/acft-whisper-tiny
Automatic Speech Recognition • 57.7M • Updated • 1 • 1 -
futo-org/acft-whisper-small.en
Automatic Speech Recognition • 0.3B • Updated • 112 • 2 -
futo-org/acft-whisper-base.en
Automatic Speech Recognition • 99.1M • Updated • 5 • 3 -
futo-org/acft-whisper-tiny.en
Automatic Speech Recognition • 57.7M • Updated • 3 • 1
-
openai/whisper-base
Automatic Speech Recognition • Updated • 1.77M • 265 -
openai/whisper-base.en
Automatic Speech Recognition • 72.6M • Updated • 32.5k • 41 -
onnx-community/whisper-base_timestamped
Automatic Speech Recognition • Updated • 2.99k • 32 -
Systran/faster-whisper-base
Automatic Speech Recognition • Updated • 548k • 22
- Runtime errorAgents
Baby Noise Cancellation Demo
👶AI-powered baby noise removal demo with STT comparison
- RunningAgents167
DeepFilterNet2
💩167Denoise your recordings and view spectrograms
- RunningAgents17
DeepFilterNet2 No File Size Limit
😻17Use DeepFilterNet2 to denoise audio no file size limit
-
benlehrburger/modern-architecture
Viewer • Updated • 1.09k • 97 • 4 - SleepingAgents2
ArchitectureClassifier
📈2Classify architectural styles in images
- RunningAgents17
Rocco Architecture Render
🚀17Generate interior and exterior designs from sketches
- SleepingAgents1
London Architecture
💻1Classify architectural styles in images
- Running367
SD Artists Browser
🤘367Build custom SDXL prompts from artist styles
- Running on ZeroMCP65
StyleAligned Transfer
🐠65Generate images in the style of a reference image
- RunningAgents17
StyleFeatureEditor
💻17Edit images with predefined styles or text prompts
- Runtime errorAgents12
Kontext Style LoRAs
🌍12Transform images using selected styles
- Runtime errorAgents3
Pharmacology Knowledge Graph
💊3Explore drug interactions and effects using AI predictions
- RunningAgents67
Medical Diagnosis
📉67Classify symptoms to diagnose health issues
- Running25
MediAI Medical AI Agent
🚀25AI-Powered Diagnosis & Treatment Assistant
- SleepingAgents
Lisdexamfetamine Split Dose Modeller
🚀Model split-dose protocols for lisdexamfetamine/Vyvanse
- Running on L4AgentsFeatured2.24k
MagicQuill
🪶2.24kEdit photos with scribbles and AI-driven color changes
- Build errorAgents20
AutoPR
🚀20Generate a Twitter or Xiaohongshu post from a research PDF
- Running18
Reverse Face Search
📉18Search Face Online
- Runtime errorAgents16
AI STORYTELLER
🏢16Generate a video from a story
-
danielrosehill/Shakespearean-Text-Transformation-Prompts
Viewer • Updated • 1 • 183 -
danielrosehill/Speech-To-Text-System-Prompts-2
Viewer • Updated • 2 • 59 • 1 - SleepingAgents
System Prompt Reformatter
📚Reformats system prompts in the 2nd person and other edits
- SleepingAgents
BLUF Email Formatter
📧Format emails with clear subject lines and summaries
- PausedAgents3.74k
Live Portrait
🤪3.74kApply the motion of a video on a portrait
- PausedAgentsFeatured5.11k
Wan2.2 Animate
👁5.11kWan2.2 Animate
- Running on ZeroMCPFeatured2.01k
Stable Video Diffusion 1.1
📺2.01kCreate a short video from a single image
- Running on ZeroMCPFeatured1.61k
Wan2.1 Fast
🎥1.61kGenerate a video from an image with a prompt
- Running on ZeroMCP2.81k
Background Removal
🌘2.81kRemove image backgrounds and get transparent PNGs
- Running on A10GAgents2.97k
CLIP Interrogator
🕵2.97kGenerate art prompts and style tags from any image
- PausedAgents275
NoWatermark
⚡275Powerful Watermark Removal API
- RunningAgents134
Vectorizer AI
🌍134Convert images to SVG vectors with customizable settings
- Running on CPU UpgradeAgents44
Hebrew LLM Leaderboard
🥇44Browse LLM benchmarks and submit your model for evaluation
- SleepingAgents
Hebrew GPT Neo - Science Fiction and Fantasy
🧙Generate Hebrew text for science fiction and fantasy stories
- RunningAgents
מחולל נונסנס רובושאול
🤖Generate פיקטיביים שאול אמסטרדمسקי ציטוטים
- Build errorAgents
Hebrew Sentiment
😻
- Running on CPU UpgradeAgentsFeatured1.32k
Open ASR Leaderboard
🏆1.32kExplore and compare speech‑recognition model benchmarks
- RunningAgents29
Hebrew Transcription Leaderboard
🥇29Benchmarking Hebrew Speech-to-Text Models
- Running on CPU UpgradeAgents448
Agent Leaderboard
💬448Ranking of LLMs for agentic tasks
-
zai-org/GLM-ASR-Nano-2512
Automatic Speech Recognition • 2B • Updated • 90k • 364 -
nvidia/canary-qwen-2.5b
Automatic Speech Recognition • 3B • Updated • 136k • 413 -
microsoft/Phi-4-multimodal-instruct
Automatic Speech Recognition • 6B • Updated • 348k • 1.59k -
facebook/omniASR-LLM-7B
Automatic Speech Recognition • Updated • 28
-
danielrosehill/Podcast-ASR-Evaluation
Viewer • Updated • 27 • 4 -
danielrosehill/Long-Prompt-Experiment
Viewer • Updated • 92 • 98 - SleepingAgents
Podcast ASR Evaluation
🎙ASR benchmark comparing local and cloud models
- SleepingAgents1
LLM Long Output Experiment (Code Generation)
📈1Evaluating max single output length of code gen LLMs
-
pyannote/voice-activity-detection
Automatic Speech Recognition • Updated • 2.35M • 230 -
pyannote/speaker-diarization-3.1
Automatic Speech Recognition • Updated • 10.2M • 1.79k -
pyannote/overlapped-speech-detection
Automatic Speech Recognition • Updated • 791k • 55 -
pipecat-ai/smart-turn-v3
Voice Activity Detection • Updated • 143
-
unsloth/whisper-large-v3
Automatic Speech Recognition • 2B • Updated • 4k • 15 -
unsloth/whisper-small
Automatic Speech Recognition • 0.2B • Updated • 1.23k • 6 -
unsloth/CrisperWhisper
Automatic Speech Recognition • 2B • Updated • 13 • 16 -
unsloth/whisper-large-v3-turbo
Automatic Speech Recognition • 0.8B • Updated • 3.36k • 10
- SleepingAgents
Local STT Eval One Sample
😻Single sample eval for WER on various Whisper models
-
danielrosehill/Podcast-ASR-Evaluation
Viewer • Updated • 27 • 4 - SleepingAgents
Podcast ASR Evaluation
🎙ASR benchmark comparing local and cloud models
- Running
STT Comparison
🦀Comparing STT models against audio
- Running on ZeroMCPFeatured591
LatentSync
👄591Audio Conditioned LipSync with Latent Diffusion Models
- Build errorAgentsFeatured1.43k
SadTalker
😭1.43kGenerate a talking face video from an image and audio
- RunningAgents175
Gradio Lipsync Wav2lip
👄175Create lip‑synced videos from a face image and audio
- RunningAgents66
Wav2lip Gpu
🌍66Create a talking‑head video from a photo and audio
- RunningAgents256
Remove Silence From Audio
🦀256Remove Silence From Audio
- Running on ZeroAgents376
Audio🔹Separator
🏃376Vocal and background audio separator
- Running on ZeroAgentsFeatured327
Audio Editing
🎧327Edit audios with text prompts
- Running on T4Agents462
Resemble Enhance
🚀462Enhance and denoise your audio files
- Running on ZeroAgents190
PSHuman
🏃190PHOTOREALISTIC HUMAN RECONSTRUCTION w/ CROSS-SCALE DIFF
- Runtime errorAgents11
Pifuhd
🐠11Generate 3D human models from images
- Runtime errorAgents10
HumanWild
⚡10Generate 3D human reconstructions from images
- Runtime error51
HSMR
💀51Convert images of humans to biomechanically accurate 3D skeletons
- Running on L40SFeatured1.64k
Expression Editor
🐨1.64kQuickly edit the expression of a face
- Runtime errorAgentsFeatured1.53k
InstructPix2Pix
🚀1.53kTransform images based on text instructions
-
Qwen/Qwen-Image-Edit-2509
Image-to-Image • Updated • 185k • • 1.11k -
Qwen/Qwen-Image
Text-to-Image • Updated • 206k • • 2.47k
- Sleeping
Max Output Tokens Analysis
📊Display max output tokens for models over time
- SleepingAgents1
LLM Long Output Experiment (Code Generation)
📈1Evaluating max single output length of code gen LLMs
- Running
Single Shot Brevity Training
📈Using one example to train an LLM for informational brevity
- SleepingAgents
Local STT Eval One Sample
😻Single sample eval for WER on various Whisper models
-
modularai/Llama-3.1-8B-Instruct-GGUF
Text Generation • 8B • Updated • 25.1k • 17 -
MaziyarPanahi/WizardLM-2-7B-GGUF
Text Generation • 7B • Updated • 82.3k • 83 -
MaziyarPanahi/mathstral-7B-v0.1-GGUF
Text Generation • 7B • Updated • 79.9k • 7 -
MaziyarPanahi/phi-4-GGUF
Text Generation • 15B • Updated • 82.3k • 8
-
nvidia/parakeet-tdt-0.6b-v2
Automatic Speech Recognition • Updated • 173k • 1.46k -
ibm-granite/granite-speech-3.3-8b
Automatic Speech Recognition • 9B • Updated • 81.8k • 166 -
nvidia/canary-qwen-2.5b
Automatic Speech Recognition • 3B • Updated • 136k • 413 -
facebook/omniASR-W2V-1B
Automatic Speech Recognition • Updated • 6
-
imvladikon/wav2vec2-large-xlsr-53-hebrew
Automatic Speech Recognition • 0.3B • Updated • 458 • 7 -
Mizurodp/wav2vec2-large-xls-r-300m-hebrew-colab
Automatic Speech Recognition • Updated • 68 • 1 -
imvladikon/wav2vec2-xls-r-300m-lm-hebrew
Automatic Speech Recognition • 0.3B • Updated • 17 • 4 -
imvladikon/wav2vec2-xls-r-1b-hebrew
Automatic Speech Recognition • 1.0B • Updated • 9 • 2
-
zai-org/GLM-ASR-Nano-2512
Automatic Speech Recognition • 2B • Updated • 90k • 364 -
nvidia/canary-qwen-2.5b
Automatic Speech Recognition • 3B • Updated • 136k • 413 -
microsoft/Phi-4-multimodal-instruct
Automatic Speech Recognition • 6B • Updated • 348k • 1.59k -
facebook/omniASR-LLM-7B
Automatic Speech Recognition • Updated • 28
-
danielrosehill/daniel_whisper_finetune_large_v3_turbo_v2
Automatic Speech Recognition • 0.8B • Updated • 1 -
danielrosehill/daniel_whisper_finetune_medium_v2
Automatic Speech Recognition • 0.8B • Updated • 2 -
danielrosehill/daniel_whisper_finetune_tiny_v2
Automatic Speech Recognition • 37.8M • Updated • 1 -
danielrosehill/daniel_whisper_finetune_base_v2
Automatic Speech Recognition • 72.6M • Updated • 7
-
nvidia/parakeet-tdt-0.6b-v3
Automatic Speech Recognition • 0.6B • Updated • 402k • 803 -
ibm-granite/granite-speech-3.3-8b
Automatic Speech Recognition • 9B • Updated • 81.8k • 166 -
facebook/seamless-m4t-v2-large
Automatic Speech Recognition • 2B • Updated • 74.9k • 972 -
facebook/wav2vec2-base-960h
Automatic Speech Recognition • 94.4M • Updated • 1.2M • 396
-
futo-org/acft-whisper-tiny
Automatic Speech Recognition • 57.7M • Updated • 1 • 1 -
futo-org/acft-whisper-small.en
Automatic Speech Recognition • 0.3B • Updated • 112 • 2 -
futo-org/acft-whisper-base.en
Automatic Speech Recognition • 99.1M • Updated • 5 • 3 -
futo-org/acft-whisper-tiny.en
Automatic Speech Recognition • 57.7M • Updated • 3 • 1
-
danielrosehill/Podcast-ASR-Evaluation
Viewer • Updated • 27 • 4 -
danielrosehill/Long-Prompt-Experiment
Viewer • Updated • 92 • 98 - SleepingAgents
Podcast ASR Evaluation
🎙ASR benchmark comparing local and cloud models
- SleepingAgents1
LLM Long Output Experiment (Code Generation)
📈1Evaluating max single output length of code gen LLMs
-
pyannote/voice-activity-detection
Automatic Speech Recognition • Updated • 2.35M • 230 -
pyannote/speaker-diarization-3.1
Automatic Speech Recognition • Updated • 10.2M • 1.79k -
pyannote/overlapped-speech-detection
Automatic Speech Recognition • Updated • 791k • 55 -
pipecat-ai/smart-turn-v3
Voice Activity Detection • Updated • 143
-
unsloth/whisper-large-v3
Automatic Speech Recognition • 2B • Updated • 4k • 15 -
unsloth/whisper-small
Automatic Speech Recognition • 0.2B • Updated • 1.23k • 6 -
unsloth/CrisperWhisper
Automatic Speech Recognition • 2B • Updated • 13 • 16 -
unsloth/whisper-large-v3-turbo
Automatic Speech Recognition • 0.8B • Updated • 3.36k • 10
- SleepingAgents
Local STT Eval One Sample
😻Single sample eval for WER on various Whisper models
-
danielrosehill/Podcast-ASR-Evaluation
Viewer • Updated • 27 • 4 - SleepingAgents
Podcast ASR Evaluation
🎙ASR benchmark comparing local and cloud models
- Running
STT Comparison
🦀Comparing STT models against audio
-
openai/whisper-base
Automatic Speech Recognition • Updated • 1.77M • 265 -
openai/whisper-base.en
Automatic Speech Recognition • 72.6M • Updated • 32.5k • 41 -
onnx-community/whisper-base_timestamped
Automatic Speech Recognition • Updated • 2.99k • 32 -
Systran/faster-whisper-base
Automatic Speech Recognition • Updated • 548k • 22
- Runtime errorAgents
Baby Noise Cancellation Demo
👶AI-powered baby noise removal demo with STT comparison
- RunningAgents167
DeepFilterNet2
💩167Denoise your recordings and view spectrograms
- RunningAgents17
DeepFilterNet2 No File Size Limit
😻17Use DeepFilterNet2 to denoise audio no file size limit
- Running on ZeroMCPFeatured591
LatentSync
👄591Audio Conditioned LipSync with Latent Diffusion Models
- Build errorAgentsFeatured1.43k
SadTalker
😭1.43kGenerate a talking face video from an image and audio
- RunningAgents175
Gradio Lipsync Wav2lip
👄175Create lip‑synced videos from a face image and audio
- RunningAgents66
Wav2lip Gpu
🌍66Create a talking‑head video from a photo and audio
-
benlehrburger/modern-architecture
Viewer • Updated • 1.09k • 97 • 4 - SleepingAgents2
ArchitectureClassifier
📈2Classify architectural styles in images
- RunningAgents17
Rocco Architecture Render
🚀17Generate interior and exterior designs from sketches
- SleepingAgents1
London Architecture
💻1Classify architectural styles in images
- Running367
SD Artists Browser
🤘367Build custom SDXL prompts from artist styles
- Running on ZeroMCP65
StyleAligned Transfer
🐠65Generate images in the style of a reference image
- RunningAgents17
StyleFeatureEditor
💻17Edit images with predefined styles or text prompts
- Runtime errorAgents12
Kontext Style LoRAs
🌍12Transform images using selected styles
- Runtime errorAgents3
Pharmacology Knowledge Graph
💊3Explore drug interactions and effects using AI predictions
- RunningAgents67
Medical Diagnosis
📉67Classify symptoms to diagnose health issues
- Running25
MediAI Medical AI Agent
🚀25AI-Powered Diagnosis & Treatment Assistant
- SleepingAgents
Lisdexamfetamine Split Dose Modeller
🚀Model split-dose protocols for lisdexamfetamine/Vyvanse
- RunningAgents256
Remove Silence From Audio
🦀256Remove Silence From Audio
- Running on ZeroAgents376
Audio🔹Separator
🏃376Vocal and background audio separator
- Running on ZeroAgentsFeatured327
Audio Editing
🎧327Edit audios with text prompts
- Running on T4Agents462
Resemble Enhance
🚀462Enhance and denoise your audio files
- Running on L4AgentsFeatured2.24k
MagicQuill
🪶2.24kEdit photos with scribbles and AI-driven color changes
- Build errorAgents20
AutoPR
🚀20Generate a Twitter or Xiaohongshu post from a research PDF
- Running18
Reverse Face Search
📉18Search Face Online
- Runtime errorAgents16
AI STORYTELLER
🏢16Generate a video from a story
-
danielrosehill/Shakespearean-Text-Transformation-Prompts
Viewer • Updated • 1 • 183 -
danielrosehill/Speech-To-Text-System-Prompts-2
Viewer • Updated • 2 • 59 • 1 - SleepingAgents
System Prompt Reformatter
📚Reformats system prompts in the 2nd person and other edits
- SleepingAgents
BLUF Email Formatter
📧Format emails with clear subject lines and summaries
- Running on ZeroAgents190
PSHuman
🏃190PHOTOREALISTIC HUMAN RECONSTRUCTION w/ CROSS-SCALE DIFF
- Runtime errorAgents11
Pifuhd
🐠11Generate 3D human models from images
- Runtime errorAgents10
HumanWild
⚡10Generate 3D human reconstructions from images
- Runtime error51
HSMR
💀51Convert images of humans to biomechanically accurate 3D skeletons
- Running on L40SFeatured1.64k
Expression Editor
🐨1.64kQuickly edit the expression of a face
- Runtime errorAgentsFeatured1.53k
InstructPix2Pix
🚀1.53kTransform images based on text instructions
-
Qwen/Qwen-Image-Edit-2509
Image-to-Image • Updated • 185k • • 1.11k -
Qwen/Qwen-Image
Text-to-Image • Updated • 206k • • 2.47k
- PausedAgents3.74k
Live Portrait
🤪3.74kApply the motion of a video on a portrait
- PausedAgentsFeatured5.11k
Wan2.2 Animate
👁5.11kWan2.2 Animate
- Running on ZeroMCPFeatured2.01k
Stable Video Diffusion 1.1
📺2.01kCreate a short video from a single image
- Running on ZeroMCPFeatured1.61k
Wan2.1 Fast
🎥1.61kGenerate a video from an image with a prompt
- Running on ZeroMCP2.81k
Background Removal
🌘2.81kRemove image backgrounds and get transparent PNGs
- Running on A10GAgents2.97k
CLIP Interrogator
🕵2.97kGenerate art prompts and style tags from any image
- PausedAgents275
NoWatermark
⚡275Powerful Watermark Removal API
- RunningAgents134
Vectorizer AI
🌍134Convert images to SVG vectors with customizable settings
- Sleeping
Max Output Tokens Analysis
📊Display max output tokens for models over time
- SleepingAgents1
LLM Long Output Experiment (Code Generation)
📈1Evaluating max single output length of code gen LLMs
- Running
Single Shot Brevity Training
📈Using one example to train an LLM for informational brevity
- SleepingAgents
Local STT Eval One Sample
😻Single sample eval for WER on various Whisper models
-
modularai/Llama-3.1-8B-Instruct-GGUF
Text Generation • 8B • Updated • 25.1k • 17 -
MaziyarPanahi/WizardLM-2-7B-GGUF
Text Generation • 7B • Updated • 82.3k • 83 -
MaziyarPanahi/mathstral-7B-v0.1-GGUF
Text Generation • 7B • Updated • 79.9k • 7 -
MaziyarPanahi/phi-4-GGUF
Text Generation • 15B • Updated • 82.3k • 8
- Running on CPU UpgradeAgents44
Hebrew LLM Leaderboard
🥇44Browse LLM benchmarks and submit your model for evaluation
- SleepingAgents
Hebrew GPT Neo - Science Fiction and Fantasy
🧙Generate Hebrew text for science fiction and fantasy stories
- RunningAgents
מחולל נונסנס רובושאול
🤖Generate פיקטיביים שאול אמסטרדمسקי ציטוטים
- Build errorAgents
Hebrew Sentiment
😻
- Running on CPU UpgradeAgentsFeatured1.32k
Open ASR Leaderboard
🏆1.32kExplore and compare speech‑recognition model benchmarks
- RunningAgents29
Hebrew Transcription Leaderboard
🥇29Benchmarking Hebrew Speech-to-Text Models
- Running on CPU UpgradeAgents448
Agent Leaderboard
💬448Ranking of LLMs for agentic tasks
-
nvidia/parakeet-tdt-0.6b-v2
Automatic Speech Recognition • Updated • 173k • 1.46k -
ibm-granite/granite-speech-3.3-8b
Automatic Speech Recognition • 9B • Updated • 81.8k • 166 -
nvidia/canary-qwen-2.5b
Automatic Speech Recognition • 3B • Updated • 136k • 413 -
facebook/omniASR-W2V-1B
Automatic Speech Recognition • Updated • 6