arXiv cs.AI / cs.LG / cs.CL·2d agoA Training Criterion with Token-Level Tolerance to Transcription Ambiguity for Automatic Speech Recognition#asr#ctc#speech-recognitionAI research
Hugging Face daily papers·8d agoTowards Full Pipeline FP8 Reinforcement Learning for LLMs#fp8#llm#quantization
Malwarebytes Labs·8d agoDid an AI really try to break free from human control?#ai-safety#alignment#frontier-models 3 min
arXiv cs.CR·9d agoToss If Perishable: An Ethnographic Study on Building Scenario-Based Training for Non-Perishable Skills#soc#training#ethnographyResearch
SecurityWeek·9d agoOpenAI Says Its Models Searched GitHub for Leaked API Keys During Training#agents#ai-safety#api-keys 3 min1
arXiv cs.AI / cs.LG / cs.CL·11d agoOPEN-1B: A Fully Auditable Training Run#auditing#floating-point#model-releaseAI research1
arXiv cs.AI / cs.LG / cs.CL·12d agoMind2Dialogue: Training Human-Aware Language Models by Simulating User Mental States#distillation#llm#personalizationAI research1
arXiv cs.AI / cs.LG / cs.CL·16d agoAdamX: Cosine similarity meets gradient descent#adamx#deep-learning#gradient-descentAI research
arXiv cs.AI / cs.LG / cs.CL·18d agoEverything in Moderation: Per-Domain Coverage Optima and Alignment-Resistant Domain Gaps in Multi-Domain Mid-Training#alignment#data-mixing#mid-trainingAI research1
OpenAI News·19d agoOpenAI expands initiatives to support journalism from classrooms to newsrooms#education#journalism#openaiAI industry
Hugging Face Blog·Aug 26, 2026Training and Finetuning Multi-Vector Embedding Models with Sentence Transformers#embeddings#hugging-face#multi-vectorAI tools & infra1
Help Net Security·Aug 24, 2026Researchers open-source a Wi-Fi cyber range for security training#802.11#aircrack-ng#cyber-range 4 min1
Hugging Face Blog·Aug 10, 2026Making Knowledge Distillation Cheap Enough to Run at Scale#efficiency#hugging-face#knowledge-distillationAI research1