Infosecurity Magazine·5d ago highExperts Alarmed Over Gyazo’s Breach of 490 Million Metadata Records#breach#data-leak#exif 2 sources in the wild 3 min
Hugging Face daily papers·6d agoAll-in-One Multilingual Scene Text Recognition with Script-aware Mixture-of-Experts#scene-text-recognition#mixture-of-experts#multilingual
DataBreaches.net·8d agoHHS’ Office for Civil Rights Settles HIPAA Investigation of Ambry Genetics for Security Rule Violations#ambry-genetics#compliance#genetic-dataPolicy & legal
MarkTechPost·8d agoJina AI Releases jina-ocr-v1: A 3.4B MoE Document Parser With Built-In Speculative Decoding for Low-Budget GPUs#deepseek-ocr#document-parsing#jina-ai 5 min1
Hugging Face daily papers·10d agoWeVisDoc: From Coverage to Capability for Robust End-to-End Document Parsing#data-synthesis#document-parsing#ocr2
arXiv cs.AI / cs.LG / cs.CL·11d agoTables Decoded: DELTA for Structure, TARQA for Understanding#document-intelligence#multilingual#ocrAI research1
MarkTechPost·17d agoLandingAI Releases Agentic Document Extraction Gen2 with DPT-3 Pro and DPT-3 Verity#developer-tools#document-intelligence#dpt-3-pro 4 min
MarkTechPost·19d agoReducto Releases r-1: A Single Pass Document Parsing Model That Cuts Errors 20% at 1 Cent Per Page#api#document-parsing#ocr 4 min1
Hugging Face daily papers·24d agoHow Far Can Synthetic Data Take Thai OCR?#document-ai#low-resource#ocr