Milos T
Smartground
·
AI & ML interests
None yet
Recent Activity
updated a collection 4 days ago
Papers updated a collection 4 days ago
Papers updated a collection 4 days ago
PapersOrganizations
None yet
Automata
General
Speako
-
ibm-granite/granite-speech-3.2-8b
Automatic Speech Recognition • 8B • Updated • 62.3k • 88 -
ByteDance/MegaTTS3
Text-to-Speech • Updated • 473 • 418 - SleepingAgents2
Demo
🚀2Transcribe and translate audio/video files with speaker diarization
-
nvidia/audio-flamingo-3-hf
Audio-Text-to-Text • 8B • Updated • 115k • 193
Imagen
Data
Code
Papers
-
Demystifying Agent Skills: Why They Work-Until They Don't
Paper • 2608.14036 • Published • 166 -
Harness the Memory: A Holistic Evaluation of Memory Substrates in Memory Agents
Paper • 2608.15008 • Published • 15 -
Oracle Agent Memory as an Enterprise Memory Substrate for Long-Horizon AI Agents
Paper • 2607.13157 • Published -
SelfMem: Self-Optimizing Memory for AI Agents
Paper • 2607.03726 • Published
Edge
Audio
-
stabilityai/stable-audio-open-small
Text-to-Audio • 0.5B • Updated • 1.98k • 273 - RunningFeatured95
ONNX Model Explorer
🔍95Explore and visualize ONNX machine‑learning models
-
microsoft/VibeVoice-1.5B
Text-to-Speech • 3B • Updated • 125k • 2.47k -
nvidia/audio-flamingo-3
Audio-Text-to-Text • Updated • 260 • 157
Play-Ground
OCR
Spatial
Multimode
Bookmark
Papers
-
Demystifying Agent Skills: Why They Work-Until They Don't
Paper • 2608.14036 • Published • 166 -
Harness the Memory: A Holistic Evaluation of Memory Substrates in Memory Agents
Paper • 2608.15008 • Published • 15 -
Oracle Agent Memory as an Enterprise Memory Substrate for Long-Horizon AI Agents
Paper • 2607.13157 • Published -
SelfMem: Self-Optimizing Memory for AI Agents
Paper • 2607.03726 • Published
Automata
Edge
General
Audio
-
stabilityai/stable-audio-open-small
Text-to-Audio • 0.5B • Updated • 1.98k • 273 - RunningFeatured95
ONNX Model Explorer
🔍95Explore and visualize ONNX machine‑learning models
-
microsoft/VibeVoice-1.5B
Text-to-Speech • 3B • Updated • 125k • 2.47k -
nvidia/audio-flamingo-3
Audio-Text-to-Text • Updated • 260 • 157
Speako
-
ibm-granite/granite-speech-3.2-8b
Automatic Speech Recognition • 8B • Updated • 62.3k • 88 -
ByteDance/MegaTTS3
Text-to-Speech • Updated • 473 • 418 - SleepingAgents2
Demo
🚀2Transcribe and translate audio/video files with speaker diarization
-
nvidia/audio-flamingo-3-hf
Audio-Text-to-Text • 8B • Updated • 115k • 193
Play-Ground
Imagen
OCR
Data
Spatial
Code
Multimode