TechniqueRLHF / Alignment8 recent entries11 Aug 2026Explaining, Verifying, and Aligning Semantic Hierarchies in Vision-Language Model EmbeddingsarXiv:2603.26798v2 Announce Type: replace-cross Abstract: Vision-language model (VLM) encoders such as CLIP enable strong retrieval and zero-shot classification in a shared image-text embedding space,→11 Aug 2026EFFEKT: Efficient Federated Knowledge Transfer to Foundation ModelsarXiv:2608.08138v1 Announce Type: new Abstract: Recent data protection laws have accelerated the adoption of Federated Learning (FL) for privacy-preserving decentralized training. Nevertheless, increa
TechniqueRAG8 recent entries31 Jul 2026ChronoMem: Version Control and Semantic Rollback for Large Language Model Agent MemoryarXiv:2607.27773v1 Announce Type: new Abstract: LLM agents increasingly rely on long-term memory to support multi-session interaction and personalization. However, existing agent memory systems are de→31 Jul 2026AgentMap: Joint Equivalence and Subsumption Discovery for Ontology MatchingarXiv:2607.27130v1 Announce Type: new Abstract: Ontology matching (OM) has traditionally been formulated as either equivalence discovery or subsumption matching. The existing OM systems identify only →3 Aug 2026I benchmarked classic vector RAG vs Google's new OKF format vs both combined — same corpus, same 7 questions, all local (Ollama + ChromaDB)Google Cloud published OKF (Open Knowledge Format) on June 12th — a spec for storing curated knowledge as a directory of markdown files with YAML frontmatter. One concept per file, linked to each othe→5 Aug 2026ArtECulture: Benchmarking Culture-Conditioned Visual Emotion Understanding in Multimodal Large Language ModelsarXiv:2608.03358v1 Announce Type: new Abstract: Existing visual emotion understanding methods typically ignore cultural variations in emotional perception. We introduce culture-conditioned visual emot→10 Aug 2026Factorized Hypothesis Search for Evidence-to-Taxonomy RetrievalarXiv:2608.06614v1 Announce Type: cross Abstract: Large-taxonomy retrieval often assumes that the input already expresses the target concept. In many settings, however, the input is indirect evidence,→11 Aug 2026SAKE: Structured Agentic Knowledge Extrapolation for Complex LLM Reasoning via Reinforcement LearningarXiv:2505.15062v5 Announce Type: replace-cross Abstract: Knowledge extrapolation is the process of inferring novel information by combining and extending existing knowledge that is explicitly availab→11 Aug 2026Forgotten History or Test-of-Time? Retrospect and Prospect on RAG from an IR PerspectivearXiv:2608.08445v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) is widely regarded as a novel paradigm born from the limitations of large language models (LLMs)--a mechanism to gr→12 Aug 2026TRACE: Trustworthy Retrieval-Augmented Conversational EnginearXiv:2608.10176v1 Announce Type: new Abstract: Public service chatbots are expected to deliver recommendations from an underlying public service directory, while also making sure that the recommendat
TechniqueAgents8 recent entries11 Aug 2026Agentic Stage-One Stellarator Optimization: Autonomous Multi-Objective Search for Finite-Beta EquilibriaarXiv:2608.01344v2 Announce Type: replace Abstract: Stage-one stellarator design searches a high-dimensional family of three-dimensional plasma boundaries and fixed-boundary MHD equilibria for configu→12 Aug 2026V-FiLLM: Verified Financial LLM Reasoning BenchmarkarXiv:2608.11047v1 Announce Type: new Abstract: While existing benchmarks have made substantial progress in evaluating LLMs across STEM domains, financial reasoning over structured data remains compar→12 Aug 2026TRACE: Trustworthy Retrieval-Augmented Conversational EnginearXiv:2608.10176v1 Announce Type: new Abstract: Public service chatbots are expected to deliver recommendations from an underlying public service directory, while also making sure that the recommendat→12 Aug 2026Toward a Theory of Value in AI AlignmentarXiv:2608.10327v1 Announce Type: new Abstract: Can AI systems be aligned to human values? The popularization of large language models (LLMs) and multi-modal foundation models has seen a rise in harms→12 Aug 2026ReLTEx: Reliable LLM-based Taxonomy ExpansionarXiv:2608.10970v1 Announce Type: cross Abstract: Recent advances in Large Language Models (LLMs) have demonstrated strong capabilities in generating semantically relevant concepts and relations, maki→12 Aug 2026Inferential Capability Does Not Determine Legal ScopearXiv:2608.10601v1 Announce Type: cross Abstract: Two instruments of EU digital law place inference at their centre and mean different things by it. Article 3(1) of the AI Act uses the capability to i→12 Aug 2026HSSBench: Benchmarking Humanities and Social Sciences Ability for Multimodal Large Language ModelsarXiv:2506.03922v4 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated significant potential to advance a broad range of domains. However, current benchma→12 Aug 2026Hierarchical Compositionality for An Assistive AI AgentarXiv:2608.10330v1 Announce Type: new Abstract: AI agents are increasingly being developed to assist humans in various applications, and Large Language Models and other deep network architectures are
TechniqueFine-tuning8 recent entries10 Aug 2026CASA: Classification Augmented with Safety Attention for Robust Multimodal AlignmentarXiv:2604.00310v2 Announce Type: replace-cross Abstract: Multimodal large-language models (MLLMs) often experience degraded safety alignment when harmful queries exploit cross-modal interactions. Mod→10 Aug 2026A foundation-model approach to pediatric headache classification from rs-fMRIarXiv:2608.07287v1 Announce Type: new Abstract: Headache is the most common neurological disorder in children and substantially affects quality of life. We investigated whether resting-state functiona→11 Aug 2026SAKE: Structured Agentic Knowledge Extrapolation for Complex LLM Reasoning via Reinforcement LearningarXiv:2505.15062v5 Announce Type: replace-cross Abstract: Knowledge extrapolation is the process of inferring novel information by combining and extending existing knowledge that is explicitly availab→11 Aug 2026MiniMax-H3: ~38 GB less VRAM with Runtime LoRA Bypass — DoRA Dynamic LoRA Loader v1.0.39GitHub: https://github.com/xmarre/ComfyUI-DoRA-Dynamic-LoRA-Loader Release v1.0.39: https://github.com/xmarre/ComfyUI-DoRA-Dynamic-LoRA-Loader/releases/tag/v1.0.39 Also available through ComfyUI Manag→11 Aug 2026FlowErase-OPD: Multi-Concept Erasure via Anchored On-Policy Distillation in Flow Matching ModelsarXiv:2608.07620v1 Announce Type: new Abstract: Recent advances in flow matching models have substantially improved the quality of text-to-image generation, but have also raised increasing safety conc→11 Aug 2026Exploring LLM Capabilities for Situational Understanding and COLREG compliance on real-world maritime navigation scenariosarXiv:2608.08281v1 Announce Type: new Abstract: Recently, Large Language Models (LLMs) have shown considerable capability for situational understanding, reasoning, and decision making in different dom→11 Aug 2026EFFEKT: Efficient Federated Knowledge Transfer to Foundation ModelsarXiv:2608.08138v1 Announce Type: new Abstract: Recent data protection laws have accelerated the adoption of Federated Learning (FL) for privacy-preserving decentralized training. Nevertheless, increa→12 Aug 2026V-FiLLM: Verified Financial LLM Reasoning BenchmarkarXiv:2608.11047v1 Announce Type: new Abstract: While existing benchmarks have made substantial progress in evaluating LLMs across STEM domains, financial reasoning over structured data remains compar
TechniqueMultimodal8 recent entries11 Aug 2026Is CLIP Cross-Eyed? Revealing and Mitigating Center Bias in the CLIP FamilyarXiv:2604.05971v2 Announce Type: replace-cross Abstract: Recent research has shown that contrastive vision-language models such as CLIP often lack fine-grained understanding of visual content. While →11 Aug 2026I gave DeepSeek V4 Flash basic vision by training a 40M connector on 100K examplesI wanted to find out whether a huge text-only MoE could be given basic vision without retraining the language model itself. The short answer is yes. I froze DeepSeek V4 Flash and a 417M-parameter Moon→11 Aug 2026Explaining, Verifying, and Aligning Semantic Hierarchies in Vision-Language Model EmbeddingsarXiv:2603.26798v2 Announce Type: replace-cross Abstract: Vision-language model (VLM) encoders such as CLIP enable strong retrieval and zero-shot classification in a shared image-text embedding space,→11 Aug 2026CFD-Guided Detection of Concept Drift in Multimodal Physiologic SignalsarXiv:2608.07759v1 Announce Type: cross Abstract: Cardiovascular AI models can classify clean elec- trocardiogram (ECG) signals, but real wearable signals change because of motion, breathing, posture,→12 Aug 2026SeFaR: Semantic Feature-aware Robustness Testing of Deep Neural NetworksarXiv:2608.10289v1 Announce Type: new Abstract: Deep neural networks are increasingly deployed in safety-critical domains as perception modules, where failures are often caused due to rare and under-r→12 Aug 2026ReCBM: Uncertainty-Gated Relational Reasoning for Concept Bottleneck ModelsarXiv:2608.10004v1 Announce Type: new Abstract: Concept Bottleneck Models (CBMs) provide an interpretable framework by grounding predictions in human-understandable concepts, enabling semantic inspect→12 Aug 2026Logit Lens Supervision for Patch-Level Explanations in Vision-Language ModelsarXiv:2602.01530v2 Announce Type: replace Abstract: Modern autoregressive Vision-Language Models (VLMs) can generate fluent answers while their visual-token representations become weakly tied to the i→12 Aug 2026HSSBench: Benchmarking Humanities and Social Sciences Ability for Multimodal Large Language ModelsarXiv:2506.03922v4 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated significant potential to advance a broad range of domains. However, current benchma
TechniqueSafety8 recent entries11 Aug 2026Beyond Hazard Resemblance: Contrastive Event Adjudication for Training-Free Video Anomaly DetectionarXiv:2608.09908v1 Announce Type: new Abstract: Video anomaly detection (VAD) aims to identify and temporally localize abnormal events in videos. Supervised methods learn anomaly decision boundaries f→12 Aug 2026Toward a Theory of Value in AI AlignmentarXiv:2608.10327v1 Announce Type: new Abstract: Can AI systems be aligned to human values? The popularization of large language models (LLMs) and multi-modal foundation models has seen a rise in harms→12 Aug 2026The Illusion of Cross-Lingual Safety in Low-Resource LanguagesarXiv:2608.11146v1 Announce Type: new Abstract: Safety alignment in large language models (LLMs) is largely developed in English, assuming these safeguards generalize across multilingual settings. How→12 Aug 2026SeFaR: Semantic Feature-aware Robustness Testing of Deep Neural NetworksarXiv:2608.10289v1 Announce Type: new Abstract: Deep neural networks are increasingly deployed in safety-critical domains as perception modules, where failures are often caused due to rare and under-r→12 Aug 2026Introspective Attention Modulation for Safe Text-to-Image GenerationarXiv:2607.14945v2 Announce Type: replace Abstract: State-of-the-art flow based text-to-image (T2I) models exhibit remarkable generative abilities but remain vulnerable to producing unsafe content. Pr→12 Aug 2026Injecting Hallucinations in Autonomous Vehicles: A Component-Agnostic Safety Evaluation FrameworkarXiv:2510.07749v2 Announce Type: replace Abstract: Perception failures in autonomous vehicles (AV) remain a major safety concern because they are the basis for many accidents. To study how these fail→12 Aug 2026Evaluation-Conditioned Training: Teaching Models to Generalize to Stronger Oversight RegimesarXiv:2608.10209v1 Announce Type: new Abstract: Feedback signals used to train Large Language Models (LLMs) are the primary driver of their behavior and our main lever for instilling alignment with hu→12 Aug 2026A Convolutional Layer Activation Dimensionality Reduction for Out-of-Distribution and Adversarial Attack Detection MethodsarXiv:2608.10203v1 Announce Type: new Abstract: Despite the success of convolutional neural networks in image classification tasks and their general application in multi-modal models, their susceptibi