TechniqueRLHF / Alignment8 recent entries7 Aug 2026Cautious Context Steering for Language Model PersonalizationarXiv:2608.05813v1 Announce Type: new Abstract: Personalizing language models (LMs) to individual user preferences is essential for aligning responses with diverse goals and backgrounds. Existing meth→10 Aug 2026InstanceSplat: Instance-Aware Feed-Forward 3D Gaussian Splatting for Scene UnderstandingarXiv:2608.07144v1 Announce Type: new Abstract: Feed-forward 3D Gaussian Splatting (3DGS) enables efficient and generalizable 3D reconstruction, but current feed-forward 3DGS methods for scene underst
TechniqueRAG8 recent entries29 Jul 2026HVM-GraphRAG: Holistic-View Multimodal Graph Retrieval-Augmented Generation on Complex DocumentarXiv:2607.24861v1 Announce Type: cross Abstract: Question answering (QA) over complex documents requires models to retrieve and integrate evidence distributed across distant document regions and moda→31 Jul 2026Bridging the Gap in Ophthalmic AI: MM-Retinal-Reason Dataset and OphthaReason Model toward Dynamic Multimodal ReasoningarXiv:2508.16129v3 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have recently demonstrated remarkable reasoning abilities with reinforcement learning paradigm. Although se→3 Aug 2026Copy Less, Ground More: Overcoming Repetitive Copying in Long-Context Reasoning via Evidence-Aware Reinforcement LearningarXiv:2607.19345v2 Announce Type: replace-cross Abstract: Large language models that generate step-by-step reasoning traces have achieved strong performance on complex tasks, and extending them to lon→6 Aug 2026The RAIL Principles for Neurosymbolic AI: Reasoning, Assurances, Interfacing and LearningarXiv:2608.04285v1 Announce Type: new Abstract: Neurosymbolic AI systems that integrate machine learning and symbolic reasoning are rapidly gaining attention. They complement the data-intensive statis→7 Aug 2026Mapping Similarity Spaces across Embedding Models with Synthetic Query ProbingarXiv:2608.05857v1 Announce Type: new Abstract: Retrieval-Augmented Generation systems rely on similarity scores to retrieve relevant content, yet scores are not directly comparable across embedding m→7 Aug 2026100% Local RAG Without Internet and Without OllamaBuild a 100% offline fast Retrieval Augmented Generation (RAG) system that runs without an internet connection, without cloud APIs, without OpenAI/Ollama Published a video where you can build a fully →10 Aug 2026Improving Attributed Long-form Question Answering with Intent AwarenessarXiv:2603.27435v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly being used to generate comprehensive, knowledge-intensive reports. However, while these models a→12 Aug 2026ThinkRetrieve: Retrieval-Augmented Reasoning Traces for Test-Time ScalingarXiv:2608.10928v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) improve performance by allocating additional inference-time compute to generate extended chain-of-thought reasoning. Howev
TechniqueAgents8 recent entries6 Aug 2026Build visibility for Codex on Amazon Bedrock with OpenTelemetry and Amazon CloudWatchAs engineering teams adopt coding agents like Codex, leaders need visibility into adoption, consumption, and reliability. This post shows how to route Codex OpenTelemetry metrics through a local colle→11 Aug 2026MaxModShift: Model Privacy via Designed ShiftsarXiv:2608.09328v1 Announce Type: new Abstract: Model learning by an eavesdropper is treated as an estimation problem in a federated environment. The Fisher Information Matrix for the eavesdropper's e→11 Aug 2026I’ve been thinking about how agents can learn inside world models for years. We decided to scale up our RSI Lab to bridge recursive self-imp…I’ve been thinking about how agents can learn inside world models for years. We decided to scale up our RSI Lab to bridge recursive self-improvement with physical AI and robotics. We are looking for f→11 Aug 2026From Trajectories to Evidence: Auditable Experimental Records for Industrial Research AgentsarXiv:2608.05235v1 Announce Type: cross Abstract: Research agents increasingly conduct multi-round machine-learning experiments in industrial recommendation settings and retain the resulting trajector→11 Aug 2026Experience-Sensitive Game Learning: A Behavioral Study of Humans and Language AgentsarXiv:2608.07490v1 Announce Type: cross Abstract: Large language model agents are increasingly evaluated through games, but most benchmarks emphasize final outcomes rather than how players learn from →11 Aug 2026DSLE: A Learning Environment for Dark Souls Boss EncountersarXiv:2608.09902v1 Announce Type: new Abstract: We introduce the Dark Souls Learning Environment (DSLE), a containerized platform that presents all 22 boss encounters of Dark Souls: Remastered as game→12 Aug 2026Long-Horizon AI Research for Grothendieck Constant: A Case Study in Human-AI Mathematical CollaborationarXiv:2608.11195v1 Announce Type: new Abstract: AI agents are increasingly used in mathematics research, but it is often unclear how to use them effectively. Towards this, we present an extensive case→12 Aug 2026Live now: our Memory & Continual Learning Track from AI Engineer World's Fair 2026. Thesis: we scaled intelligence and got the world's smart…Live now: our Memory & Continual Learning Track from AI Engineer World's Fair 2026. Thesis: we scaled intelligence and got the world's smartest novice. https://www.youtube.com/watch?v=iqloyWCGYQQ&list
TechniqueFine-tuning8 recent entries11 Aug 2026TS-Mob: Social and Geographical-Aware Time Series Foundation-Model Framework for Human Mobility PredictionarXiv:2507.00945v2 Announce Type: replace Abstract: Short-term forecasting of aggregated human mobility flows supports urban planning, intelligent transportation systems, and emergency response, yet e→11 Aug 2026TLDChoiceNet: Quantitatively Choosing a Transfer Learning DatasetarXiv:2608.09091v1 Announce Type: cross Abstract: Transfer learning is particularly useful in settings with limited training data, and within image classification it is common to transfer learn upon m→11 Aug 2026SignLlama: Enhancing Gloss-free Sign Language Translation by Prioritizing Visual Features for LLMsarXiv:2608.09006v1 Announce Type: cross Abstract: Large Language Models (LLMs) have achieved remarkable success across a wide range of tasks. However, fine-tuning LLMs for Gloss-Free Sign Language Tra→11 Aug 2026SafeQL: Search-based Refinement for Safe and Efficient LLM-based Text-to-SQLarXiv:2608.09260v1 Announce Type: cross Abstract: Large language models (LLMs) have advanced Text-to-SQL by enabling natural language interfaces to databases without task-specific fine-tuning. However→11 Aug 2026Multi-Granular Node Pruning for Causal Circuit DiscoveryarXiv:2512.10903v3 Announce Type: replace Abstract: Circuit discovery aims to identify minimal subnetworks that are responsible for specific behaviors in large language models (LLMs). Existing approac→12 Aug 2026Link-adaptive digital twin for robust physical-layer modeling in hybrid-amplified ultra-wideband optical networksarXiv:2608.10517v1 Announce Type: cross Abstract: Accurate physical-layer modeling is increasingly essential for reliable ultra-wideband operation and capacity optimization, especially under the inten→12 Aug 2026Finding the Signal in the Spam: Jointly Learning Rewards and Worker Reliability from Pairwise ComparisonsarXiv:2608.10045v1 Announce Type: cross Abstract: The problem of learning from pairwise comparisons has been widely studied across many domains such as recommendation systems, social choice, and more →12 Aug 2026Can Bayesian Optimization Efficiently Find a Strong Single Expert in Neural Thickets?arXiv:2608.10867v1 Announce Type: new Abstract: Gradient-free post-training has emerged as a compelling alternative to gradient-based optimization for large language models (LLMs), but existing approa
TechniqueMultimodal8 recent entries11 Aug 2026Let Geometry GUIDE: Layer-wise Unrolling of Geometric Priors in Multimodal LLMsarXiv:2604.05695v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable progress in 2D visual tasks but still struggle to understand physical space in rea→11 Aug 2026LAD-COD: Language-Aligned Dense Perception for Camouflaged Object DetectionarXiv:2608.07941v1 Announce Type: new Abstract: Camouflaged object detection (COD) aims to segment objects that exhibit high visual similarity to their surroundings, which reduces foreground-backgroun→11 Aug 2026From Objectives to What Models Learn: A Landau Theory of Invariant LearningarXiv:2608.09396v1 Announce Type: new Abstract: Invariant learning seeks representations that remain predictive across environments, yet the behavior of its objectives along the regularization path is→11 Aug 2026Did the Grid Erase the Event? EndoClock for Auditing Medical World-Model PipelinesarXiv:2608.09266v1 Announce Type: new Abstract: Medical world models commonly learn from multimodal recordings synchronized onto a fixed-rate grid. This preprocessing resamples each native stream onto→12 Aug 2026Transformer Geometry Observatory TGO-IV: Developmental Topology ObservatoryarXiv:2608.09997v1 Announce Type: cross Abstract: Transformers have had a profound impact on the world of language processing and computer vision. As efforts to answer the million-dollar question of `→12 Aug 2026ReCBM: Uncertainty-Gated Relational Reasoning for Concept Bottleneck ModelsarXiv:2608.10004v1 Announce Type: new Abstract: Concept Bottleneck Models (CBMs) provide an interpretable framework by grounding predictions in human-understandable concepts, enabling semantic inspect→12 Aug 2026Rationale-Guided Learning for Multimodal Emotion RecognitionarXiv:2608.10448v1 Announce Type: new Abstract: Multimodal emotion recognition in conversation (MERC) requires understanding complex interactions between verbal and non-verbal cues. However, most exis→12 Aug 2026Once Poisoned, Arbitrarily Controlled: A Programmable Backdoor in VLMsarXiv:2608.10959v1 Announce Type: new Abstract: Existing vision-language model (VLM) backdoors are usually treated as static vulnerabilities: one-to-one and N-to-N attacks bind one or more triggers to
TechniqueSafety8 recent entries2 Jul 2026From World Models to World Action Models: A Concise Tutorial for RoboticsarXiv:2607.00836v1 Announce Type: cross Abstract: World models are increasingly used in embodied intelligence and generative simulation, yet their scope remains ambiguous across communities. This tuto→5 Jul 2026Wiki Lint Report — 2026-07-05Automated lint: 51 errors, 15 warnings, 3 info→15 Jul 2026Wiki Lint Report — 2026-07-15Automated lint: 26 errors, 6728 warnings, 3 info→19 Jul 2026Wiki Lint Report — 2026-07-19Automated lint: 20 errors, 8743 warnings, 3 info→24 Jul 2026Is Your Safe Controller Actually Safe? A Critical Review of CBF Tautologies and Hidden AssumptionsarXiv:2603.06954v2 Announce Type: replace Abstract: This tutorial provides a critical review of the practical application of Control Barrier Functions (CBFs) in robotic safety. While the theoretical f→25 Jul 2026Old Coder Needs help with New AI Development and wants to get up to speed to understand it all.Hi Guys, I'm an old coder and DBA that has been in the field for almost 40 years. More and more the jobs I was doing for work are being taken over by AI and the need for my type of work is diminishing→28 Jul 2026An Unofficial FastLAS Tutorial: A Programmer's GuidearXiv:2607.23557v1 Announce Type: cross Abstract: FastLAS is a scalable system for Inductive Logic Programming (ILP): you give it some background knowledge, a language bias, and a set of examples, and→11 Aug 2026DSLE: A Learning Environment for Dark Souls Boss EncountersarXiv:2608.09902v1 Announce Type: new Abstract: We introduce the Dark Souls Learning Environment (DSLE), a containerized platform that presents all 22 boss encounters of Dark Souls: Remastered as game