Post
3086
🇹🇷 **A 1B Turkish OCR model vs. Baidu OCR.**
We ran both models on the **same Turkish enterprise documents**.
The result:
**Werea-DocOCR-1B → 99.9**
**Baidu Unlimited-OCR → 47.9**
Same documents.
Same evaluation.
And Werea-DocOCR is only **1B parameters**.
It was built specifically for difficult Turkish enterprise documents:
📄 invoices
📑 contracts
🏦 bank receipts
💼 payroll
🚗 vehicle documents
📋 SGK-style tables
📱 scanned & phone-captured documents
But benchmarks aren't enough.
**Give me a Turkish document that you think will break it.**
We'll test the hardest ones and publish the failures.
🤗 Model:
Werea-co/Werea-DocOCR-1B
📚 Dataset:
Werea-co/werea-tr-doc-ocr-enterprise-v2
🇹🇷 Built in Türkiye. Open on Hugging Face.
#OCR #DocumentAI #HuggingFace #TurkishAI
We ran both models on the **same Turkish enterprise documents**.
The result:
**Werea-DocOCR-1B → 99.9**
**Baidu Unlimited-OCR → 47.9**
Same documents.
Same evaluation.
And Werea-DocOCR is only **1B parameters**.
It was built specifically for difficult Turkish enterprise documents:
📄 invoices
📑 contracts
🏦 bank receipts
💼 payroll
🚗 vehicle documents
📋 SGK-style tables
📱 scanned & phone-captured documents
But benchmarks aren't enough.
**Give me a Turkish document that you think will break it.**
We'll test the hardest ones and publish the failures.
🤗 Model:
Werea-co/Werea-DocOCR-1B
📚 Dataset:
Werea-co/werea-tr-doc-ocr-enterprise-v2
🇹🇷 Built in Türkiye. Open on Hugging Face.
#OCR #DocumentAI #HuggingFace #TurkishAI