arXiv cs.AI / cs.LG / cs.CL·15d agoAttention Quantization for Tabular Foundation Models#fp8#inference-efficiency#quantizationAI research1