CompanyMeta3 recent entries16 Apr 2026Language steering in latent space to mitigate unintended code-switchingarXiv:2510.13849v3 Announce Type: replace Abstract: Multilingual Large Language Models (LLMs) often exhibit hallucinations such as unintended code-switching, reducing reliability in downstream tasks. →20 May 2026Contrastive Reasoning Alignment: Reinforcement Learning from Hidden RepresentationsarXiv:2603.17305v2 Announce Type: replace Abstract: We propose CRAFT, a red-teaming alignment framework that leverages model reasoning capabilities and hidden representations to improve robustness aga
CompanyMistral1 recent entries11 Aug 2026Do All LLMs Know When They're Being Harmful? A Reproducibility Study of Latent-Space Safety Probes Across Model FamiliesarXiv:2608.08029v1 Announce Type: cross Abstract: Khatri et al. (2026) [DOI: 10.1109/DSN-W70714.2026.00027] show that lightweight MLP probes on final-layer activations of a single 8B model (LLaMA-3.1-
CompanyNVIDIA2 recent entries5 Jun 2026Predict and Reconstruct: Joint Objectives for Self-Supervised Language Representation LearningarXiv:2606.05173v1 Announce Type: new Abstract: Masked language modelling (MLM) has been the dominant pre-training objective for text encoders since BERT, yet it encourages representations that are st→6 Aug 2026AI-based single-shot structured-light depth reconstruction for real-time laparoscopic surgical guidancearXiv:2608.05109v1 Announce Type: cross Abstract: Significance. Accurate intraoperative depth perception is important for autonomous and semi-autonomous robotic laparoscopic surgery. Conventional frin