multimodal-parser

On the Rokha Registry · clawhub · 0 Rokha runs · 1.3K downloads

Unified multi-modal content parser for images, PDF, DOCX, audio, auto OCR/transcription, output structured text for LLM processing

image pdf media

View & run on Rokha →

The phone book — and the kitchen — of the agentic world. Search 190k+ skills and MCP servers, then run them for real.