Hugging Face daily papers·8d agoTowards Full Pipeline FP8 Reinforcement Learning for LLMs#fp8#llm#quantization
Hugging Face daily papers·13d agoVC-Attention: Value Smoothing and Softmax Casting for Low-bit Attention#diffusion-transformers#fp8#gpu-kernels1
arXiv cs.AI / cs.LG / cs.CL·15d agoAttention Quantization for Tabular Foundation Models#fp8#inference-efficiency#quantizationAI research1
Hacker News · AI·16d agoDeepSeek v4.1 Flash Uncensored#deepseek#deepseek-v4.1-flash#fp8Model release1
Hugging Face daily papers·16d agoZGCM-1: A Fully Open and Extremely Efficient Foundation Model for Math and Agentic Search#7b#agentic#foundation-model1
Hacker News · security·17d agoTraining a 3.8B LLM to 0.384 CORE for $998 – Hugo Vergnes#core-benchmark#fp8#little-lmAI research 15 min1
Hugging Face trending models·27d agodealignai/GLM-5.3-CYBERSECURITY-FP8 — new model trending #13 on Hugging Face#cybersecurity-llm#fp8#glm-5.3Model release 5 min1