You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
DeepSeek-OCR-2-Unlimited-OCR is an advanced, experimental visual document processing and open-ended text localization dashboard. This application establishes a unified interface that allows users to swap between two premier vision-language document models: deepseek-ai/DeepSeek-OCR-2 and baidu/Unlimited-OCR.
A Gradio-based demonstration application for the Tencent HunyuanOCR model, focused on optical character recognition (OCR) tasks such as text detection, extraction, and coordinate formatting from images. Users can upload images, customize prompts (e.g., for Chinese/English text).
DeepSeek-OCR-experimental is an advanced, multi-purpose visual document intelligence and object localization sandbox. Powered by the unredacted prithivMLmods/DeepSeek-OCR-Latest-BF16.I64-v2.0 architecture, this suite is designed to deliver highly accurate, structure-aware image text extractions.
🖥️ Utilize DeepSeek-OCR-2 to effortlessly execute advanced OCR tasks, converting documents to markdown and extracting text through an intuitive web app.
This repo is used to map the Controller value to the Mesh Vertex by using an Artificial Neural Network. A machine learning–based facial deformation system that detects facial landmarks and applies controlled transformations to modify facial expressions or geometry in real time.