AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
5,189 results
Agents

MapAgent: An Industrial-Grade Agentic Framework for City-scale Lane-level Map Generation

DGX agent

arXiv:2606.04513v1 Announce Type: new Abstract: Lane-level maps are critical infrastructure for autonomous driving and lane-level navigation, yet constructing and maintaining standardized lane network

agentsarxiv-cs-ai
4 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Measuring Model Robustness via Fisher Information: Spectral Bounds, Theoretical Guarantees, and Practical Algorithms

DGX agent

arXiv:2606.04767v1 Announce Type: cross Abstract: The robustness of deep neural networks is crucial for safety-critical deployments, yet existing evaluation methods are often attack-dependent and lack

safetyarxiv-cs-cv
4 Jun 2026
Research

MuCO: Generative Peptide Cyclization Empowered by Multi-stage Conformation Optimization

DGX agent

arXiv:2602.11189v2 Announce Type: replace-cross Abstract: Modeling peptide cyclization is critical for the virtual screening of candidate peptides with desirable physical and pharmaceutical properties

researcharxiv-cs-ai
4 Jun 2026
Safety

Neetyabhas: A Framework for Uncertainty-Aware Public Policy Optimization in Rational Agent-Based Models

DGX agent

arXiv:2606.04562v1 Announce Type: new Abstract: Purpose The WHO's COVID-19 non-pharmaceutical interventions (e.g., lockdowns, vaccinations) effectively curb transmission but impose heavy economic stra

safetyarxiv-cs-ai
4 Jun 2026
Safety

Policy Gradient for Continuous-Time Robust Markov Decision Processes

DGX agent

arXiv:2606.04335v1 Announce Type: new Abstract: The framework of robust Markov decision processes (RMDPs) allows the design of reinforcement learning agents that satisfy performance guarantees under w

safetyarxiv-cs-lg
4 Jun 2026
Applications

Reconciling Causality and Non-Equilibrium Thermodynamics with Hamiltonian Causal Models

DGX agent

arXiv:2606.04822v1 Announce Type: new Abstract: Causal modeling of physical temporal phenomena must handle interventions that act along trajectories, nonstationary induced laws, path-dependent effects

applicationsarxiv-cs-lg
4 Jun 2026
Research

Reconstructing Unobservable Temperature Fields via Simulation-Aided Intelligent Sensing

DGX agent

arXiv:2606.04582v1 Announce Type: cross Abstract: Real-time monitoring of the temperature distribution within components and sub-structures is a challenging topic in many systems due to restrictions o

researcharxiv-cs-lg
4 Jun 2026
Local Ai

Recover-LoRA for Aggressive Quantization: Reclaiming Accuracy in 2-Bit Language Models via Low-Rank Adaptation with Knowledge Distillation on Synthetic Data

DGX agent

arXiv:2606.04238v1 Announce Type: cross Abstract: Aggressive weight quantization to 2-bit precision offers substantial throughput and memory gains for large language model (LLM) inference, but typical

local-aiarxiv-cs-ai
4 Jun 2026
Safety

Reinforcement Learning from Rich Feedback with Distributional DAgger

DGX agent

arXiv:2606.05152v1 Announce Type: cross Abstract: Reasoning models have advanced rapidly, but the dominant reinforcement learning from verifiable rewards (RLVR) recipe remains surprisingly narrow: sam

safetyarxiv-cs-ai
4 Jun 2026
Safety

Rethinking Continual Experience Internalization for Self-Evolving LLM Agents

DGX agent

arXiv:2606.04703v1 Announce Type: new Abstract: Experience internalization converts contextual experience from past interactions into reusable parametric capability, offering a promising path toward c

safetyarxiv-cs-cl
4 Jun 2026
Safety

The Invisible Lottery: How Subtle Cues Steer Algorithm Choice in LLM Code Generation

DGX agent

arXiv:2606.04057v1 Announce Type: cross Abstract: Large language models (LLMs) now generate substantial production code, often for tasks with multiple valid algorithmic solutions. Incidental prompt cu

safetyarxiv-cs-ai
4 Jun 2026
Safety

Transferable Multi-Bit Watermarking Across Frozen Diffusion Models via Latent Consistency Bridges

DGX agent

arXiv:2603.20304v2 Announce Type: replace Abstract: As generative AI advances, global governance frameworks increasingly mandate verifiable content provenance. However, existing watermarking technique

safetyarxiv-cs-cv
4 Jun 2026
Safety

U-Net-Accelerated Quality-Diversity Optimization for Climate-Adaptive Urban Layouts

DGX agent

arXiv:2606.04658v1 Announce Type: cross Abstract: Optimizing urban layouts for climate adaptation requires balancing building density with cold-air ventilation. Because physics-based climate simulatio

safetyarxiv-cs-lg
4 Jun 2026
Research

Uncovering Insights of Compound Flooding with Data-Driven AI

DGX agent

arXiv:2506.04281v2 Announce Type: replace Abstract: Compound flooding, driven by nonlinear interactions between multiple hydrometeorological factors, poses a significant challenge to hazard prevention

researcharxiv-cs-lg
4 Jun 2026
Model Releases

What If Prompt Injection Never Left? Exploring Cross-Session Stored Prompt Injection in Agentic Systems

DGX agent

arXiv:2606.04425v1 Announce Type: cross Abstract: Modern agentic systems transform LLMs from session-bounded assistants into stateful systems that persist and evolve shared world state across sessions

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Agent Skills for Large Language Models: Architecture, Acquisition, Security, and the Path Forward

DGX agent

arXiv:2602.12430v4 Announce Type: replace-cross Abstract: The transition from monolithic language models to modular, skill-equipped agents marks a defining shift in how large language models (LLMs) ar

model-releasesarxiv-cs-ai
3 Jun 2026
Research

AlphaEval: A Comprehensive and Efficient Evaluation Framework for Formula Alpha Mining

DGX agent

arXiv:2508.13174v2 Announce Type: replace Abstract: Formula alpha mining, which generates predictive signals from financial data, is critical for quantitative investment. Although various algorithmic

researcharxiv-cs-ai
3 Jun 2026
Model Releases

An Asymptotic Theory of Chain-of-Thought in In-Context Learning

DGX agent

arXiv:2606.03217v1 Announce Type: cross Abstract: Chain-of-thought (CoT) reasoning has become a widely used mechanism for eliciting multi-step reasoning in large language models by generating intermed

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

AUDITFLOW: Executable Symbolic Environments for Structured Financial Reporting Verification

DGX agent

arXiv:2606.03031v1 Announce Type: new Abstract: Structured financial audit verification is difficult for language-model agents because correctness depends on structured evidence rather than text alone

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Beyond Encoder Accumulation: Measuring Encoder Roles in Multi-Encoder VLMs

DGX agent

arXiv:2606.03879v1 Announce Type: cross Abstract: As foundation models scale toward fusing more heterogeneous visual streams, understanding how diverse encoders interact under joint training becomes a

model-releasesarxiv-cs-ai
3 Jun 2026
Agents

Closed-Loop Molecular Design with Calibrated Deference

DGX agent

arXiv:2606.02618v1 Announce Type: cross Abstract: We present Cognitive Loop via In-Situ Optimization (CLIO), an agent that couples a continuously-updated belief-state graph with a recursive plan-then-

agentsarxiv-cs-ai
3 Jun 2026
Safety

CP-Agent: Context-Aware Multimodal Reasoning for Cellular Morphological Profiling under Chemical Perturbations

DGX agent

arXiv:2606.03435v1 Announce Type: new Abstract: Cell Painting combines multiplexed fluorescent staining, high-content imaging, and quantitative analysis to generate high-dimensional phenotypic readout

safetyarxiv-cs-ai
3 Jun 2026
Research

CropCraft: A Procedural World Generator for Robotic Simulation of Agricultural Tasks

DGX agent

arXiv:2511.02417v2 Announce Type: replace Abstract: The adoption of agroecological practices in modern agriculture requires robotic systems capable of operating in highly diverse and complex field env

researcharxiv-cs-cv
3 Jun 2026
Research

DiffUNet^2: Bidirectional Prediction, Probabilistic Generation and Collaborative Visual Discovery for Scientific Data

DGX agent

arXiv:2606.03926v1 Announce Type: cross Abstract: Modeling temporal evolution is important to analyzing and reasoning about scientific phenomena, yet most machine learning methods provide deterministi

researcharxiv-cs-lg
3 Jun 2026
Local Ai

DMT-CBT: Longitudinal Therapeutic State Modeling for CBT Counseling

DGX agent

arXiv:2606.03132v1 Announce Type: new Abstract: Large language models (LLMs) have shown growing potential for Cognitive Behavioral Therapy (CBT) counseling. However, most existing approaches still for

local-aiarxiv-cs-cl
3 Jun 2026
Research

Don't Gamble, GAMBLe: An Analytical Framework for AI-Driven Research Systems

DGX agent

arXiv:2606.02863v1 Announce Type: new Abstract: AI-Driven Research Systems (ADRS) -- systems coupling LLMs with automated evaluation to discover algorithms, proofs, and designs -- are being optimized

researcharxiv-cs-ai
3 Jun 2026
Safety

Easy-to-Use Shielding for Reinforcement Learning

DGX agent

arXiv:2606.03804v1 Announce Type: new Abstract: Safe exploration is a key challenge in Reinforcement Learning (RL) that aims to prevent agents from making harmful decisions while exploring their envir

safetyarxiv-cs-lg
3 Jun 2026
Model Releases

EURO-5K: When Does Domain Pretraining Matter? Benchmarking Transformers for EU Reporting Obligation Extraction

DGX agent

arXiv:2606.02971v1 Announce Type: new Abstract: Extracting reporting obligations from EU legislation is critical for assessing and reducing regulatory reporting burden. However, distinguishing reporti

model-releasesarxiv-cs-cl
3 Jun 2026
Agents

EvoDS: Self-Evolving Autonomous Data Science Agent with Skill Learning and Context Management

DGX agent

arXiv:2606.03841v1 Announce Type: new Abstract: Recent progress in Large Language Model (LLM) agents has enabled promising advances in automated data science. However, existing approaches remain funda

agentsarxiv-cs-ai
3 Jun 2026
Model Releases

GTBench: A Curriculum-Grounded Benchmark for Evaluating LLMs as Mathematical Research Assistants in Graph Theory

DGX agent

arXiv:2606.03144v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as self-study assistants in technical disciplines, yet their reliability as mathematical reasoning as

model-releasesarxiv-cs-ai
3 Jun 2026
Research

Hierarchies of Calibration: Classification meets Regression

DGX agent

arXiv:2606.03245v1 Announce Type: cross Abstract: Concepts of calibration formalize the compatibility between probabilistic predictions and the respective outcomes. In a nutshell, the outcomes ought t

researcharxiv-cs-lg
3 Jun 2026
Safety

LAP: An Agent-to-Instrument Protocol for Autonomous Science

DGX agent

arXiv:2606.03755v1 Announce Type: new Abstract: Autonomous science is moving from demonstration to infrastructure. Large language model agents now plan experiments, and self-driving laboratories execu

safetyarxiv-cs-ai
3 Jun 2026
Research

Learning Multi-Scale Hypergraph for High-Order Brain Connectivity Analysis

DGX agent

arXiv:2606.03310v1 Announce Type: cross Abstract: Understanding complex interactions between brain regions is critical for early neurodegenerative disease classification such as Alzheimer's Disease (A

researcharxiv-cs-ai
3 Jun 2026
Model Releases

MLSkip: Data Skipping for ML Filters via Lightweight Metadata

DGX agent

arXiv:2606.03946v1 Announce Type: cross Abstract: Database vendors recently released AI functions that can be used in filter predicates. As such functions often rely on costly, black-box ML models, th

model-releasesarxiv-cs-lg
3 Jun 2026
Agents

MUSE: A Unified Agentic Harness for MLLMs

DGX agent

arXiv:2606.03005v1 Announce Type: cross Abstract: Despite rapid progress, multimodal large language models (MLLMs) still fail on tasks that humans solve effortlessly, such as navigating a grid maze fr

agentsarxiv-cs-ai
3 Jun 2026
Agents

OpenAgenet/OAN: Open Infrastructure for Trusted Agent Interconnection

DGX agent

arXiv:2606.03161v1 Announce Type: cross Abstract: OpenAgenet, abbreviated as OAN, is an open infrastructure project for trusted Agent interconnection. It addresses a problem that becomes visible when

agentsarxiv-cs-ai
3 Jun 2026
Research

Optimizing Neuro-Fuzzy and Colonial Competition Algorithms for Skin Cancer Diagnosis in Dermatoscopic Images

DGX agent

arXiv:2505.08886v2 Announce Type: replace Abstract: The rising incidence of skin cancer, coupled with limited public awareness and a shortfall in clinical expertise, underscores an urgent need for adv

researcharxiv-cs-cv
3 Jun 2026
Model Releases

Plan2Map: A Multimodal Benchmark for Document-Grounded Geospatial Boundary Reconstruction from Planning Records

DGX agent

arXiv:2606.02747v1 Announce Type: cross Abstract: Planning records define restrictions over geographic areas, but their source documents often provide only indirect spatial evidence rather than machin

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

PyraMathBench: Evaluating and Improving Mathematical Capability in Large Language Models

DGX agent

arXiv:2606.03858v1 Announce Type: new Abstract: Despite the pivotal role of numerical reasoning as the cornerstone of mathematical capabilities in large language models (LLMs) across applications, few

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Reasoning Structure of Large Language Models

DGX agent

arXiv:2606.03883v1 Announce Type: new Abstract: Large reasoning models (LRMs) are often evaluated using metrics such as final-answer accuracy or token count. However, identical scores on these metrics

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

SimuScene: Simulation-Ready Compositional 3D Scene Reconstruction from a Single Image

DGX agent

arXiv:2606.03994v1 Announce Type: new Abstract: Reconstructing interactive, simulation-ready 3D scenes from a single image is a critical bottleneck for robotic manipulation. While recent single-image

safetyarxiv-cs-cv
3 Jun 2026
Agents

SPADE: Sketch-guided Path Planning Augmented with Diffusion Experts

DGX agent

arXiv:2606.03512v1 Announce Type: cross Abstract: Path planning is essential for Autonomous Mobile Robots (AMRs). Conventional methods for incorporating human preferences into planning typically rely

agentsarxiv-cs-ai
3 Jun 2026
Model Releases

Testing LLM Arithmetic Reasoning Generalization with Automatic Numeric-Remapping Attacks

DGX agent

arXiv:2606.03606v1 Announce Type: cross Abstract: Large language models achieve strong performance on arithmetic reasoning benchmarks, and one common response to arithmetic brittleness is to delegate

model-releasesarxiv-cs-ai
3 Jun 2026
Research

The Unsampled Truth: Psychometrics in SLMs Measure Prompt Artifacts, Not Psychological Constructs

DGX agent

arXiv:2606.03357v1 Announce Type: cross Abstract: When prompting SLMs for psychometric assessments, researchers assume the outputs reflect semantic reasoning. We evaluate this premise across 13 open-w

researcharxiv-cs-ai
3 Jun 2026
Safety

Towards a Science of AI Agent Reliability

DGX agent

arXiv:2602.16666v3 Announce Type: replace Abstract: AI agents are increasingly deployed to execute important tasks. While rising accuracy scores on standard benchmarks suggest rapid progress, many age

safetyarxiv-cs-ai
3 Jun 2026
Research

Tracking Urban Atmospheric Pollutants using Sentinel-5P Satellite Data

DGX agent

arXiv:2606.02592v1 Announce Type: cross Abstract: Urban nitrogen dioxide (NO_2) is a key indicator of combustion-related air pollution and exhibits strong spatial and temporal variability in cities. T

researcharxiv-cs-ai
3 Jun 2026
Model Releases

TriEval: A Resource-Efficient Pipeline for LLM Bias, Toxicity, and Truthfulness Assessment

DGX agent

arXiv:2606.03036v1 Announce Type: new Abstract: LLMs have evolved from basic chatbots to the backbone of the AI ecosystem, now widely used in healthcare, schools, and government services. The domain-w

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

TSQAgent: Rating Time Series Data Quality via Dedicated Agentic Reasoning

DGX agent

arXiv:2606.03629v1 Announce Type: new Abstract: Assessing the quality of time series (TS) data is fundamental yet inherently challenging due to the multifaceted nature of quality dimensions. Recently,

model-releasesarxiv-cs-ai
3 Jun 2026
← Previous
1…7677787980…109
Next →