baidu/Qianfan-OCR
Image-Text-to-Text • 5B • Updated • 284k • 1.19k
World-first embodied AI world model
Generate HTML or React code from your web app description
Small and powerful reasoning LLM that runs in your browser
Scalable and Versatile 3D Generation from images
FireRed-Image-Edit × Qwen-Image-Edit-Rapid (Transformers)
Real-time video captioning in your browser
generate a video from an image with a text prompt