AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “papers”

GridTimelineEvolution
12,265 results
8 Jun 2026

Self-evolving LLM agents with in-distribution Optimization

SafetyDGX agent

arXiv:2606.07367v1 Announce Type: new Abstract: Large Language Models (LLMs) have recently emerged as powerful controllers for interactive agents in complex environments, yet training them to perform

Spline Policy: A Structured Representation for Robot Policies

Model ReleasesDGX agent

arXiv:2606.07386v1 Announce Type: new Abstract: Modern imitation-learning policies for robot manipulation often represent actions as fixed-resolution action chunks, which are simple and effective but

Stream3D-VLM: Online 3D Spatial Understanding with Incremental Geometry Priors

Model ReleasesDGX agent

arXiv:2606.06891v1 Announce Type: new Abstract: Despite advances in 3D scene understanding, existing 3D Large Multimodal Models operate in offline settings, requiring complete scene observations or pr

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SWE-Explore: Benchmarking How Coding Agents Explore Repositories

Model ReleasesDGX agent

arXiv:2606.07297v1 Announce Type: cross Abstract: Repository-level coding benchmarks such as SWE-bench have driven a rapid surge in the capabilities of coding agents. Yet they usually treat coding tas

SWE-IF: Aligning Code Evaluation with Human Preference

ResearchDGX agent

arXiv:2510.07315v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have catalyzed vibe coding, where users leverage LLMs to generate and iteratively refine code through natural lan

Synthics: Synthetic Physics-like Datasets for Machine Learning

ApplicationsDGX agent

arXiv:2606.06724v1 Announce Type: new Abstract: Representative data is fundamental in machine learning, as limited data hinders generalisation. Collecting sufficient real-world samples is often infeas

TA-RAG: Tone-Aware Retrieval-Augmented Generation for Peer-Support Health Communication

ResearchDGX agent

arXiv:2606.06794v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) successfully grounds large language model (LLM) outputs in trusted documents, but factual grounding alone is insuff

Task Editing for Generalizable 3D Visuomotor Policy Learning

Local AiDGX agent

arXiv:2606.07012v1 Announce Type: new Abstract: 3D visuomotor policies offer a promising direction for complex robotic manipulation, as depth maps and point clouds provide rich geometric information f

The Capacity of Information-Theoretic Secure Aggregation in Federated Learning

Local AiDGX agent

arXiv:2606.07277v1 Announce Type: cross Abstract: Secure aggregation allows a server to aggregate users' local updates while preserving update privacy. Existing information-theoretic problems typicall

The Effect of Training Task Diversity on In-Context Learning through the Lens of Low-Dimensional Subspaces

ResearchDGX agent

arXiv:2606.06814v1 Announce Type: cross Abstract: The transformer's emergent ability to perform in-context learning (ICL) has sparked a wide range of studies designed to understand its underlying mech

The Matrix idea of keeping humans as batteries is obviously weird... we would be more useful as dice. LLMs default to very similar kinds of …

ApplicationsDGX agent

The Matrix idea of keeping humans as batteries is obviously weird... we would be more useful as dice. LLMs default to very similar kinds of arguments & structure, and even different LLMs seem to colla

The Proxy Benders Decomposition

ResearchDGX agent

arXiv:2606.07403v1 Announce Type: cross Abstract: Benders decomposition is a fundamental framework for solving large-scale mixed-integer optimization problems with complicating variables that, when fi

The Three-Ring Architecture: Governing Agents in the Era of On-Platform Organisations

AgentsDGX agent

arXiv:2606.07119v1 Announce Type: cross Abstract: The current phase of enterprise AI deployment faces a structural failure: organisations are acquiring agentic capability without the infrastructure to

The Utility and Complexity of in- and out-of-Distribution Machine Unlearning

ResearchDGX agent

arXiv:2412.09119v3 Announce Type: replace Abstract: Machine unlearning, the process of selectively removing data from trained models, is increasingly crucial for addressing privacy concerns and knowle

TOPSIS-RAD: Ranking According to Desires

ResearchDGX agent

arXiv:2606.07253v1 Announce Type: new Abstract: Traditional TOPSIS derives its reference points -- the Positive Ideal Solution (PIS) and Negative Ideal Solution (NIS) -- from the observed alternative

Towards Tight Bounds for Streaming Attention

Model ReleasesDGX agent

arXiv:2606.07205v1 Announce Type: cross Abstract: The attention mechanism is a cornerstone of modern transformer architectures. However, its expressive power comes at the cost of quadratic runtime and

VALUEFLOW: Toward Pluralistic and Steerable Value-based Alignment in Large Language Models

SafetyDGX agent

arXiv:2602.03160v2 Announce Type: replace Abstract: Aligning Large Language Models (LLMs) with the diverse spectrum of human values remains a central challenge: preference-based methods often fail to

What Do People Actually Want From AI? Mapping Preference Plurality

SafetyDGX agent

arXiv:2606.06674v1 Announce Type: new Abstract: Large Language Models (LLMs) are often fine-tuned through Reinforcement Learning from Human Feedback (RLHF) to align with people's preferences and value

When Does Multi-Agent Collaboration Help? An Entropy Perspective

AgentsDGX agent

arXiv:2602.04234v6 Announce Type: cross Abstract: Multi-agent systems (MAS) have emerged as a prominent paradigm for leveraging large language models (LLMs) to tackle complex tasks. However, the mecha

Your UnEmbedding Matrix is Secretly a Feature Lens for Text Embeddings

ResearchDGX agent

arXiv:2606.07502v1 Announce Type: new Abstract: Large language models exhibit impressive zero-shot capabilities across a wide range of downstream tasks. However, they struggle to function as off-the-s

7 Jun 2026

Wiki Lint Report — 2026-06-07

SynthesesDGX agent

Automated lint: 47 errors, 12 warnings, 3 info

One of the more interesting takes on positive alignment that have recently come out-it’s long and interesting, combining philosophy and trai…

SafetyDGX agent

One of the more interesting takes on positive alignment that have recently come out-it’s long and interesting, combining philosophy and training setups (eg reward proposals), and worth a read. What ha

6 Jun 2026

2-Step Agent: A Framework for the Interaction of a Decision Maker with AI Decision Support

AgentsDGX agent

arXiv:2602.21889v2 Announce Type: replace Abstract: Predictions from ML models support human decision making in several fields, including high-stakes ones such as healthcare and the judiciary. Yet, we

A Framework for Measuring Appropriate Reliance on Set-Valued AI Advice

ResearchDGX agent

arXiv:2606.06081v1 Announce Type: new Abstract: Appropriate reliance on AI advice has become a central research theme in human-AI collaboration. Existing frameworks have focused exclusively on point p

A Motivational Architecture for Conversational AGI

AgentsDGX agent

arXiv:2606.05411v1 Announce Type: new Abstract: Motivational architectures in cognitive AI have largely been designed for physical agents regulating bodily needs. Conversational agents operate in a di

A Pre-Registered Causal Partition of Self-Consistency Elicitation and Reward Design in RLVR

SafetyDGX agent

arXiv:2606.05932v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR) improves reasoning even when the reward signal is spurious -- assigning credit to the group-plural

Adapting Diffusion Language Models for Lossless Pixel-Level Image Transmission

ResearchDGX agent

arXiv:2606.06273v1 Announce Type: cross Abstract: Lossless pixel-level image transmission is a fundamental regime beyond semantic communications, because exact recovery requires both accurate symbol p

Agent-Orchestrated Adaptive RAG: A Comparative Study on Structured and Multi-Hop Retrieval

Model ReleasesDGX agent

arXiv:2606.05658v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) enhances Large Language Models (LLMs) by grounding their responses in external knowledge, but conventional pipeli

AIS-Based Vessel Trajectory Prediction Using Memory-Augmented Neural Networks

ResearchDGX agent

arXiv:2606.06311v1 Announce Type: new Abstract: Accurate vessel trajectory prediction is essential for safe and efficient maritime operations, enabling collision avoidance and supporting route optimiz

An Improved CNN-LSTM Based Intrusion Detection System for IoT Networks

ResearchDGX agent

arXiv:2606.05776v1 Announce Type: cross Abstract: With the rapid proliferation of IoT devices, security concerns have dramatically escalated and intrusion detection systems have become critical for pr

Assessing the Geographic Diversity of AI's Platial Representations in Image Generation

SafetyDGX agent

arXiv:2606.05188v1 Announce Type: cross Abstract: (Gen)AI diversity is not merely an ethical issue. From the perspective of geographic information science (GIScience), it could be interpreted as a fun

Balancing Image Compression and Generation with Bootstrapped Tokenization

Local AiDGX agent

arXiv:2606.05552v1 Announce Type: cross Abstract: Despite progress in image tokenization, standard methods encode redundant information by mixing all granularities within each token, thus redundancy p

Beyond Means: Topological Causal Effects under Persistent-Homology Ignorability

ResearchDGX agent

arXiv:2603.14169v2 Announce Type: replace-cross Abstract: Average treatment effects (ATE) and conditional average treatment effects (CATE) are foundational causal estimands, but they target changes in

Beyond Similarity: Trustworthy Memory Search for Personal AI Agents

AgentsDGX agent

arXiv:2606.06054v1 Announce Type: new Abstract: Personal AI agents increasingly rely on long-term memory to provide persistent personalization across sessions. However, existing memory pipelines are l

Can AI Refute Economic Theory? Evidence from Beyond the Knowledge Cutoff

Model ReleasesDGX agent

arXiv:2606.05383v1 Announce Type: cross Abstract: Can artificial intelligence (AI) refute economic theory? I document experiments in which I asked several AI models (Gemini, Refine, Claude, and ChatGP

Can LLMs Write Correct TLA+ Specifications? Evaluating Natural-Language-to-TLA+ Generation

Model ReleasesDGX agent

arXiv:2606.05792v1 Announce Type: new Abstract: TLA+ has supported industrial verification at companies such as Amazon and Microsoft, yet writing correct TLA+ specifications from natural language stil

CangLing-KnowFlow: A Unified Knowledge-and-Flow-fused Agent for Comprehensive Remote Sensing Applications

Model ReleasesDGX agent

arXiv:2512.15231v3 Announce Type: replace Abstract: The automated and intelligent processing of massive remote sensing (RS) datasets is critical in Earth observation (EO). Existing automated systems a

CausalPOI: Spatio-Temporal Graph-Based Causal Modeling for Cold-Start POI Check-in Forecasting

ApplicationsDGX agent

arXiv:2606.05413v1 Announce Type: cross Abstract: As urban environments continue to evolve rapidly, accurately modeling the dynamic behaviour of Points of Interest is essential for supporting data-dri

Comprehensive and Reliable Feature Attribution for Diverse Modalities and Models via Frequency-Domain Insights

SafetyDGX agent

arXiv:2411.18343v3 Announce Type: replace-cross Abstract: Personalized Federal learning(PFL) allows clients to cooperatively train a personalized model without disclosing their private dataset. Howeve

// Continual Learning Bench // One of the research areas with lots of investments is continual learning. While there are many efforts, there…

TutorialsDGX agent

// Continual Learning Bench // One of the research areas with lots of investments is continual learning. While there are many efforts, there is very little progress in measuring it. So the big questio

Cross-Epoch Adaptive Rollout Optimization for RL Post-Training

Model ReleasesDGX agent

arXiv:2606.05606v1 Announce Type: cross Abstract: LLM post-training often relies on reinforcement learning methods that sample multiple rollouts per prompt, yet most existing approaches use a fixed ro

Data Flow Control: Data Safety Policies for AI Agents

Model ReleasesDGX agent

arXiv:2606.05679v1 Announce Type: cross Abstract: Agents increasingly generate SQL, orchestrate pipelines, and automate data analysis on behalf of users. While recent work improves query correctness,

Detecting Perspective Shifts in Multi-agent Systems

AgentsDGX agent

arXiv:2512.05013v2 Announce Type: replace Abstract: Generative models augmented with external tools and update mechanisms (or extit{agents}) have demonstrated capabilities beyond intelligent prompting

Dimensionality Reduction for Cyberattack Classification: A Comparative Evaluation of PCA and Linear Predictive Coding

ResearchDGX agent

arXiv:2606.05584v1 Announce Type: cross Abstract: High-dimensional feature representations are widely used in machine learning-based cyberattack detection systems. However, they increase computational

Double Preconditioning (DoPr): Optimization for Test-Time Performance, not Validation Loss

SafetyDGX agent

arXiv:2606.06418v1 Announce Type: cross Abstract: Many modern applications of deep learning involve training a neural network via a one-step prediction loss (e.g., L^2 regression, cross-entropy), but

https://x.com/CAIS/status/2060031683420999844?s=20

SafetyDGX agent

https://x.com/CAIS/status/2060031683420999844?s=20 AI systems may soon help run economies, infrastructure, and military operations. But these systems are not reliably loyal or secure. An adversary can

Human oversight of agentic systems in practice: Examining the oversight work, challenges, and heuristics of developers using software agents

AgentsDGX agent

arXiv:2606.05391v1 Announce Type: cross Abstract: Autonomous software agents hold promise to increase developer productivity but make mistakes and exhibit novel failure modes, making human oversight c

Integrating Mechanistic and Data-Driven Models for Neurological Disorders through Differentiable Programming

ResearchDGX agent

arXiv:2606.06094v1 Announce Type: new Abstract: Advances in computational modeling, neuroimaging, and artificial intelligence are revolutionizing the modeling of neurological disorders for improved di

ITP-STDP: An Intrinsic-Timing Power-of-Two Learning Engine for On-Chip SNN Training

ResearchDGX agent

arXiv:2606.06159v1 Announce Type: cross Abstract: Spiking neural networks (SNNs) have the potential to emerge as the third generation of neural networks and have attracted increasing attention across

LatentWave: JEPA Pretraining for Wireless Foundation Models

SafetyDGX agent

arXiv:2606.06373v1 Announce Type: cross Abstract: Wireless foundation models have emerged as a promising alternative to building separate models for each wireless task. However, existing approaches re

LLMCodec: Adapting Video Codecs for Efficient Weight Compression of Large Language Models

Model ReleasesDGX agent

arXiv:2606.05861v1 Announce Type: cross Abstract: The rapid development of large language models(LLMs) has led to remarkable advances in natural language processing. However, the increasing scale of t

Microskill Architecture: A Modular Skill-Driven Framework for AI-Native Code Generation

AgentsDGX agent

arXiv:2606.05720v1 Announce Type: cross Abstract: Large language models and AI coding agents have reshaped software development, but the path to fully AI-native systems faces structural challenges. Ch

New research from Renmin University. Treat skill selection as a harness in its own right. If you design skill routing for personal or edge a…

Local AiDGX agent

New research from Renmin University. Treat skill selection as a harness in its own right. If you design skill routing for personal or edge agents, this work argues that the selection layer is a first-

Ontology-constrained multi-LLM scoring of hypothesis support in the predictive processing literature

Local AiDGX agent

arXiv:2606.05206v1 Announce Type: cross Abstract: Fragmentation is common in interdisciplinary fields with diverse methods and theoretical commitments. Predictive coding neuroscience is a clear exampl

Pattern Selectivity is Not Task-Causal Structure: A Cross-Architecture Mechanistic Study of Composed-Task Circuits in 1B-Class Language Models

ResearchDGX agent

arXiv:2606.05378v1 Announce Type: cross Abstract: We test whether a single screen-and-ablate recipe -- identify attention-head circuits by task-pattern selectivity, then verify by causal ablation agai

RAG Security and Privacy: Formalizing the Threat Model and Attack Surface

ApplicationsDGX agent

arXiv:2509.20324v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) is an emerging approach in natural language processing that combines large language models (LLMs) with ex

Reformulating Neural Operators in d+1 Dimensions for Embedding Evolution

ResearchDGX agent

arXiv:2505.11766v4 Announce Type: replace-cross Abstract: Neural Operators (NOs) are powerful architectures for learning mappings between function spaces. While most advances focus on refining kernel

Risk Assessment of Autonomous Driving: Integrating Technical Failures, Ethical Dilemmas, and Policy Frameworks

SafetyDGX agent

arXiv:2606.06396v1 Announce Type: new Abstract: Autonomous driving technology has the potential to reduce the large number of road traffic accidents caused by human error each year, but it also brings

Self-Commitment Latency: A Reward-Free Probe for Prompted Implicit Hacking

ResearchDGX agent

arXiv:2606.05625v1 Announce Type: new Abstract: Implicit reward hacking is hard to audit when a language model's chain of thought appears benign: a final answer may be anchored by a prompt shortcut wh

Separation Power of Equivariant Neural Networks

ResearchDGX agent

arXiv:2406.08966v3 Announce Type: replace-cross Abstract: The separation power of a machine learning model refers to its ability to distinguish between different inputs and is often used as a proxy fo

← Previous
1…113114115116117…205
Next →