AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,023 results
4 Jun 2026

MapAgent: An Industrial-Grade Agentic Framework for City-scale Lane-level Map Generation

AgentsDGX agent

arXiv:2606.04513v1 Announce Type: new Abstract: Lane-level maps are critical infrastructure for autonomous driving and lane-level navigation, yet constructing and maintaining standardized lane network

Measuring Model Robustness via Fisher Information: Spectral Bounds, Theoretical Guarantees, and Practical Algorithms

SafetyDGX agent

arXiv:2606.04767v1 Announce Type: cross Abstract: The robustness of deep neural networks is crucial for safety-critical deployments, yet existing evaluation methods are often attack-dependent and lack

MuCO: Generative Peptide Cyclization Empowered by Multi-stage Conformation Optimization

ResearchDGX agent

arXiv:2602.11189v2 Announce Type: replace-cross Abstract: Modeling peptide cyclization is critical for the virtual screening of candidate peptides with desirable physical and pharmaceutical properties

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Neetyabhas: A Framework for Uncertainty-Aware Public Policy Optimization in Rational Agent-Based Models

SafetyDGX agent

arXiv:2606.04562v1 Announce Type: new Abstract: Purpose The WHO's COVID-19 non-pharmaceutical interventions (e.g., lockdowns, vaccinations) effectively curb transmission but impose heavy economic stra

On TV Tokyo’s WBS (@wbs_tvtokyo) tonight I’ll be discussing Sakana AI’s upcoming 1T parameter model project, supported by METI’s GENIAC init…

Model ReleasesDGX agent

On TV Tokyo’s WBS (@wbs_tvtokyo) tonight I’ll be discussing Sakana AI’s upcoming 1T parameter model project, supported by METI’s GENIAC initiative. We are scaling up to build Japan’s first 1T paramete

Policy Gradient for Continuous-Time Robust Markov Decision Processes

SafetyDGX agent

arXiv:2606.04335v1 Announce Type: new Abstract: The framework of robust Markov decision processes (RMDPs) allows the design of reinforcement learning agents that satisfy performance guarantees under w

Reconciling Causality and Non-Equilibrium Thermodynamics with Hamiltonian Causal Models

ApplicationsDGX agent

arXiv:2606.04822v1 Announce Type: new Abstract: Causal modeling of physical temporal phenomena must handle interventions that act along trajectories, nonstationary induced laws, path-dependent effects

Reconstructing Unobservable Temperature Fields via Simulation-Aided Intelligent Sensing

ResearchDGX agent

arXiv:2606.04582v1 Announce Type: cross Abstract: Real-time monitoring of the temperature distribution within components and sub-structures is a challenging topic in many systems due to restrictions o

Recover-LoRA for Aggressive Quantization: Reclaiming Accuracy in 2-Bit Language Models via Low-Rank Adaptation with Knowledge Distillation on Synthetic Data

Local AiDGX agent

arXiv:2606.04238v1 Announce Type: cross Abstract: Aggressive weight quantization to 2-bit precision offers substantial throughput and memory gains for large language model (LLM) inference, but typical

Reinforcement Learning from Rich Feedback with Distributional DAgger

SafetyDGX agent

arXiv:2606.05152v1 Announce Type: cross Abstract: Reasoning models have advanced rapidly, but the dominant reinforcement learning from verifiable rewards (RLVR) recipe remains surprisingly narrow: sam

Rethinking Continual Experience Internalization for Self-Evolving LLM Agents

SafetyDGX agent

arXiv:2606.04703v1 Announce Type: new Abstract: Experience internalization converts contextual experience from past interactions into reusable parametric capability, offering a promising path toward c

The Invisible Lottery: How Subtle Cues Steer Algorithm Choice in LLM Code Generation

SafetyDGX agent

arXiv:2606.04057v1 Announce Type: cross Abstract: Large language models (LLMs) now generate substantial production code, often for tasks with multiple valid algorithmic solutions. Incidental prompt cu

Transferable Multi-Bit Watermarking Across Frozen Diffusion Models via Latent Consistency Bridges

SafetyDGX agent

arXiv:2603.20304v2 Announce Type: replace Abstract: As generative AI advances, global governance frameworks increasingly mandate verifiable content provenance. However, existing watermarking technique

U-Net-Accelerated Quality-Diversity Optimization for Climate-Adaptive Urban Layouts

SafetyDGX agent

arXiv:2606.04658v1 Announce Type: cross Abstract: Optimizing urban layouts for climate adaptation requires balancing building density with cold-air ventilation. Because physics-based climate simulatio

Uncovering Insights of Compound Flooding with Data-Driven AI

ResearchDGX agent

arXiv:2506.04281v2 Announce Type: replace Abstract: Compound flooding, driven by nonlinear interactions between multiple hydrometeorological factors, poses a significant challenge to hazard prevention

What If Prompt Injection Never Left? Exploring Cross-Session Stored Prompt Injection in Agentic Systems

Model ReleasesDGX agent

arXiv:2606.04425v1 Announce Type: cross Abstract: Modern agentic systems transform LLMs from session-bounded assistants into stateful systems that persist and evolve shared world state across sessions

3 Jun 2026

A workflow audit is no longer the best way to figure out how to use AI in your job. Despite the advice from AI labs, I'm more convinced, bec…

Model ReleasesDGX agent

A workflow audit is no longer the best way to figure out how to use AI in your job. Despite the advice from AI labs, I'm more convinced, because of AI's reasoning capabilities and long context horizon

Absolute love seeing everyone using Hermes Desktop preview! A ton of updates have already gone in and there's more to come to give yall the …

ResearchDGX agent

Nous Research announced the release of Hermes Desktop preview, expressing enthusiasm about user adoption and indicating that multiple updates have already been implemented with additional improvements

Agent Skills for Large Language Models: Architecture, Acquisition, Security, and the Path Forward

Model ReleasesDGX agent

arXiv:2602.12430v4 Announce Type: replace-cross Abstract: The transition from monolithic language models to modular, skill-equipped agents marks a defining shift in how large language models (LLMs) ar

AlphaEval: A Comprehensive and Efficient Evaluation Framework for Formula Alpha Mining

ResearchDGX agent

arXiv:2508.13174v2 Announce Type: replace Abstract: Formula alpha mining, which generates predictive signals from financial data, is critical for quantitative investment. Although various algorithmic

An Asymptotic Theory of Chain-of-Thought in In-Context Learning

Model ReleasesDGX agent

arXiv:2606.03217v1 Announce Type: cross Abstract: Chain-of-thought (CoT) reasoning has become a widely used mechanism for eliciting multi-step reasoning in large language models by generating intermed

Another banger open-source release. Miso One is an 8B text-to-speech model with real emotional range, so voiceovers carry warmth, hesitation…

Model ReleasesDGX agent

Another banger open-source release. Miso One is an 8B text-to-speech model with real emotional range, so voiceovers carry warmth, hesitation, and excitement instead of sounding flat. It's purpose-buil

AUDITFLOW: Executable Symbolic Environments for Structured Financial Reporting Verification

Model ReleasesDGX agent

arXiv:2606.03031v1 Announce Type: new Abstract: Structured financial audit verification is difficult for language-model agents because correctness depends on structured evidence rather than text alone

Beyond Encoder Accumulation: Measuring Encoder Roles in Multi-Encoder VLMs

Model ReleasesDGX agent

arXiv:2606.03879v1 Announce Type: cross Abstract: As foundation models scale toward fusing more heterogeneous visual streams, understanding how diverse encoders interact under joint training becomes a

Bring Databricks into Kiro IDE with the AI Dev Kit Power

IndustryDGX agent

This article describes how to integrate Databricks with Kiro IDE using the AI Dev Kit, enabling developers to leverage Databricks' data and AI capabilities directly within their integrated development

Bring your favorite agent into Devin Desktop using ACP: https://docs.devin.ai/desktop/acp

AgentsDGX agent

Devin Desktop now supports Agent Control Protocol (ACP), allowing users to integrate their preferred AI agents into the platform. This feature enables users to bring custom or third-party agents into

Build 2026: From observability to ROI for AI agents on any framework

AgentsDGX agent

9 min read · June 3, 2026 · Sebastian Kohlmeier Shipping an AI agent is the easy part. Keeping it accurate, safe, and accountable in production is where teams get stuck. Agents are non-deterministic.

Closed-Loop Molecular Design with Calibrated Deference

AgentsDGX agent

arXiv:2606.02618v1 Announce Type: cross Abstract: We present Cognitive Loop via In-Situ Optimization (CLIO), an agent that couples a continuously-updated belief-state graph with a recursive plan-then-

CP-Agent: Context-Aware Multimodal Reasoning for Cellular Morphological Profiling under Chemical Perturbations

SafetyDGX agent

arXiv:2606.03435v1 Announce Type: new Abstract: Cell Painting combines multiplexed fluorescent staining, high-content imaging, and quantitative analysis to generate high-dimensional phenotypic readout

CropCraft: A Procedural World Generator for Robotic Simulation of Agricultural Tasks

ResearchDGX agent

arXiv:2511.02417v2 Announce Type: replace Abstract: The adoption of agroecological practices in modern agriculture requires robotic systems capable of operating in highly diverse and complex field env

DiffUNet^2: Bidirectional Prediction, Probabilistic Generation and Collaborative Visual Discovery for Scientific Data

ResearchDGX agent

arXiv:2606.03926v1 Announce Type: cross Abstract: Modeling temporal evolution is important to analyzing and reasoning about scientific phenomena, yet most machine learning methods provide deterministi

DMT-CBT: Longitudinal Therapeutic State Modeling for CBT Counseling

Local AiDGX agent

arXiv:2606.03132v1 Announce Type: new Abstract: Large language models (LLMs) have shown growing potential for Cognitive Behavioral Therapy (CBT) counseling. However, most existing approaches still for

Don't Gamble, GAMBLe: An Analytical Framework for AI-Driven Research Systems

ResearchDGX agent

arXiv:2606.02863v1 Announce Type: new Abstract: AI-Driven Research Systems (ADRS) -- systems coupling LLMs with automated evaluation to discover algorithms, proofs, and designs -- are being optimized

Easy-to-Use Shielding for Reinforcement Learning

SafetyDGX agent

arXiv:2606.03804v1 Announce Type: new Abstract: Safe exploration is a key challenge in Reinforcement Learning (RL) that aims to prevent agents from making harmful decisions while exploring their envir

EURO-5K: When Does Domain Pretraining Matter? Benchmarking Transformers for EU Reporting Obligation Extraction

Model ReleasesDGX agent

arXiv:2606.02971v1 Announce Type: new Abstract: Extracting reporting obligations from EU legislation is critical for assessing and reducing regulatory reporting burden. However, distinguishing reporti

EvoDS: Self-Evolving Autonomous Data Science Agent with Skill Learning and Context Management

AgentsDGX agent

arXiv:2606.03841v1 Announce Type: new Abstract: Recent progress in Large Language Model (LLM) agents has enabled promising advances in automated data science. However, existing approaches remain funda

GTBench: A Curriculum-Grounded Benchmark for Evaluating LLMs as Mathematical Research Assistants in Graph Theory

Model ReleasesDGX agent

arXiv:2606.03144v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as self-study assistants in technical disciplines, yet their reliability as mathematical reasoning as

Hierarchies of Calibration: Classification meets Regression

ResearchDGX agent

arXiv:2606.03245v1 Announce Type: cross Abstract: Concepts of calibration formalize the compatibility between probabilistic predictions and the respective outcomes. In a nutshell, the outcomes ought t

I'm finally launching this agent tomorrow! Will be free, open source, and powered exclusively by open models on @togethercompute. Will also …

AgentsDGX agent

I'm finally launching this agent tomorrow! Will be free, open source, and powered exclusively by open models on @togethercompute. Will also drop a full guide on how it works! Building an agent that ca

Jerry Liu built one of the most installed pieces of AI plumbing of the last three years. Then he sat down and told me the framework era he h…

Model ReleasesDGX agent

Jerry Liu built one of the most installed pieces of AI plumbing of the last three years. Then he sat down and told me the framework era he helped create is over. The agent harness ate the abstraction

LAP: An Agent-to-Instrument Protocol for Autonomous Science

SafetyDGX agent

arXiv:2606.03755v1 Announce Type: new Abstract: Autonomous science is moving from demonstration to infrastructure. Large language model agents now plan experiments, and self-driving laboratories execu

Learning Multi-Scale Hypergraph for High-Order Brain Connectivity Analysis

ResearchDGX agent

arXiv:2606.03310v1 Announce Type: cross Abstract: Understanding complex interactions between brain regions is critical for early neurodegenerative disease classification such as Alzheimer's Disease (A

microsoft MAI tech report is a gold mine, one of the most transparent for a model at this scale. this model uses zero synthetic data or dist…

AgentsDGX agent

microsoft MAI tech report is a gold mine, one of the most transparent for a model at this scale. this model uses zero synthetic data or distillation from previous models. this means reasoning, agentic

MLSkip: Data Skipping for ML Filters via Lightweight Metadata

Model ReleasesDGX agent

arXiv:2606.03946v1 Announce Type: cross Abstract: Database vendors recently released AI functions that can be used in filter predicates. As such functions often rely on costly, black-box ML models, th

MUSE: A Unified Agentic Harness for MLLMs

AgentsDGX agent

arXiv:2606.03005v1 Announce Type: cross Abstract: Despite rapid progress, multimodal large language models (MLLMs) still fail on tasks that humans solve effortlessly, such as navigating a grid maze fr

Not to mention that having HIPAA and FERPA compliant AI systems makes thousands of students and researchers using them less risky.

ApplicationsDGX agent

HIPAA and FERPA compliant AI systems reduce institutional and legal risks when used by students and researchers by ensuring sensitive health and educational data are properly protected. Compliance wit

OpenAgenet/OAN: Open Infrastructure for Trusted Agent Interconnection

AgentsDGX agent

arXiv:2606.03161v1 Announce Type: cross Abstract: OpenAgenet, abbreviated as OAN, is an open infrastructure project for trusted Agent interconnection. It addresses a problem that becomes visible when

Optimizing Neuro-Fuzzy and Colonial Competition Algorithms for Skin Cancer Diagnosis in Dermatoscopic Images

ResearchDGX agent

arXiv:2505.08886v2 Announce Type: replace Abstract: The rising incidence of skin cancer, coupled with limited public awareness and a shortfall in clinical expertise, underscores an urgent need for adv

Plan2Map: A Multimodal Benchmark for Document-Grounded Geospatial Boundary Reconstruction from Planning Records

Model ReleasesDGX agent

arXiv:2606.02747v1 Announce Type: cross Abstract: Planning records define restrictions over geographic areas, but their source documents often provide only indirect spatial evidence rather than machin

PyraMathBench: Evaluating and Improving Mathematical Capability in Large Language Models

Model ReleasesDGX agent

arXiv:2606.03858v1 Announce Type: new Abstract: Despite the pivotal role of numerical reasoning as the cornerstone of mathematical capabilities in large language models (LLMs) across applications, few

Reasoning Structure of Large Language Models

Model ReleasesDGX agent

arXiv:2606.03883v1 Announce Type: new Abstract: Large reasoning models (LRMs) are often evaluated using metrics such as final-answer accuracy or token count. However, identical scores on these metrics

SimuScene: Simulation-Ready Compositional 3D Scene Reconstruction from a Single Image

SafetyDGX agent

arXiv:2606.03994v1 Announce Type: new Abstract: Reconstructing interactive, simulation-ready 3D scenes from a single image is a critical bottleneck for robotic manipulation. While recent single-image

SPADE: Sketch-guided Path Planning Augmented with Diffusion Experts

AgentsDGX agent

arXiv:2606.03512v1 Announce Type: cross Abstract: Path planning is essential for Autonomous Mobile Robots (AMRs). Conventional methods for incorporating human preferences into planning typically rely

Testing LLM Arithmetic Reasoning Generalization with Automatic Numeric-Remapping Attacks

Model ReleasesDGX agent

arXiv:2606.03606v1 Announce Type: cross Abstract: Large language models achieve strong performance on arithmetic reasoning benchmarks, and one common response to arithmetic brittleness is to delegate

The Hermes Web Dashboard got a major overhaul: it is now a feature-complete admin panel that you can manage entirely from your browser.

ResearchDGX agent

The Hermes Web Dashboard has been significantly redesigned to function as a comprehensive admin panel with full feature parity, allowing users to manage all operations directly through a web browser i

The next chapter in flood resilience: Open sourcing Google’s hydrology framework

ResearchDGX agent

Google has open-sourced its flood forecasting framework, which replicates operational FloodHub model training settings and reflects methodology described in a 2024 Nature paper for global ungauged flo

The Unsampled Truth: Psychometrics in SLMs Measure Prompt Artifacts, Not Psychological Constructs

ResearchDGX agent

arXiv:2606.03357v1 Announce Type: cross Abstract: When prompting SLMs for psychometric assessments, researchers assume the outputs reflect semantic reasoning. We evaluate this premise across 13 open-w

this is an interesting point in the new ted chiang piece – no one really claims that alphafold is conscious, or that sora or midjourney or d…

ResearchDGX agent

Yann LeCun discusses a point from a Ted Chiang piece about the lack of claims regarding consciousness in recent AI systems like AlphaFold, Sora, and Midjourney, suggesting skepticism about attributing

This is really intellectually dishonest. I have not been arguing that LLM token prices are increasing (though the all you can eat buffet is …

Model ReleasesDGX agent

This is really intellectually dishonest. I have not been arguing that LLM token prices are increasing (though the all you can eat buffet is over), I have been arguing the *opposite*, viz that they wil

This SkillOpt paper from Microsoft is a must-read! (bookmark it) I was a bit skeptical of the results reported in the paper when I shared it…

AgentsDGX agent

This SkillOpt paper from Microsoft is a must-read! (bookmark it) I was a bit skeptical of the results reported in the paper when I shared it a few days ago. However, I managed to integrate it into my

← Previous
1…128129130131132…168
Next →