AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,639 results
Model Releases

Causality for Tabular Data Synthesis: A High-Order Structure Causal Benchmark Framework

DGX agent

arXiv:2406.08311v3 Announce Type: replace-cross Abstract: Existing evaluations of tabular synthesis models rely primarily on low-order statistics and downstream task performance, leaving multivariate

model-releasesarxiv-cs-ai
30 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Agents

CaveAgent: Transforming LLMs into Stateful Runtime Operators

DGX agent

arXiv:2601.01569v4 Announce Type: replace Abstract: LLM-based agents are increasingly capable of complex task execution, yet current agentic systems remain constrained by text-centric paradigms that s

agentsarxiv-cs-ai
30 Jun 2026
Agents

Code Reasoning for Software Engineering Tasks: A Survey and A Call to Action

DGX agent

arXiv:2506.13932v3 Announce Type: replace-cross Abstract: The rise of large language models (LLMs) has led to dramatic improvements across a wide range of natural language tasks. Their performance on

agentsarxiv-cs-ai
30 Jun 2026
Applications

Comparing Human and Automatic Recognition of Dutch Dysarthric Continuous Speech: A Case Study

DGX agent

arXiv:2606.30237v1 Announce Type: new Abstract: In our goal to develop personalised dysarthric speech recognition (DSR) models, this study compared the recognition performances of human listeners and

applicationsarxiv-cs-cl
30 Jun 2026
Model Releases

Customized Generative AI Agent for Transportation Engineering Practice: A Development and Continued Pre-training Guideline

DGX agent

arXiv:2606.29014v1 Announce Type: new Abstract: Recent advancements in generative artificial intelligence (AI) and large language models (LLMs) have shown significant promise in automating complex rea

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Defeat Devices in AI Systems

DGX agent

arXiv:2606.28863v1 Announce Type: cross Abstract: AI systems increasingly exhibit behavior that differs systematically between evaluation and deployment contexts. Alignment faking, sandbagging, benchm

model-releasesarxiv-cs-ai
30 Jun 2026
Agents

Direct Causation in International Humanitarian Law and the Challenge of AI-Mediated Civilian Cyber Operations

DGX agent

arXiv:2606.29175v1 Announce Type: new Abstract: International humanitarian law protects civilians from direct attack unless and for such time as they take direct part in hostilities, with the ICRC's 2

agentsarxiv-cs-ai
30 Jun 2026
Model Releases

DiscoGen: Procedural Generation of Algorithm Discovery Tasks in Machine Learning

DGX agent

arXiv:2603.17863v2 Announce Type: replace-cross Abstract: Automating the development of machine learning algorithms has the potential to unlock new breakthroughs. However, our ability to improve and e

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Echoes of Human Malice in Agents: Benchmarking LLMs for Multi-Turn Online Harassment Attacks

DGX agent

arXiv:2510.14207v3 Announce Type: replace Abstract: Large Language Model (LLM) agents are powering a growing share of interactive web applications, yet remain vulnerable to misuse and harm. Prior jail

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

EPIC-EuroParl-UdS: Information-Theoretic Perspectives on Translation and Interpreting

DGX agent

arXiv:2603.09785v3 Announce Type: replace Abstract: This paper introduces an updated and combined version of the bidirectional English-German EPIC-UdS (spoken) and EuroParl-UdS (written) corpora conta

safetyarxiv-cs-cl
30 Jun 2026
Model Releases

Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions

DGX agent

arXiv:2507.05257v4 Announce Type: replace-cross Abstract: Recent benchmarks for Large Language Model (LLM) agents primarily focus on evaluating reasoning, planning, and execution capabilities, while a

model-releasesarxiv-cs-ai
30 Jun 2026
Applications

Exploring the Value of Diverse LLM Explanations in Introductory Programming

DGX agent

arXiv:2606.28882v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown the potential to generate code explanations that surpass those of peers in quality, offering promising opportu

applicationsarxiv-cs-ai
30 Jun 2026
Hardware

Featuring the man of the moment @dylan522p @shaunmmaguire in today's Training Data episode. Nobody is a more trusted industry insider to the…

DGX agent

Featuring the man of the moment @dylan522p @shaunmmaguire in today's Training Data episode. Nobody is a more trusted industry insider to the biggest infrastructure build-out in history. The story of h

hardwaresonya-huang--x
30 Jun 2026
Model Releases

fev-bench: A Realistic Benchmark for Time Series Forecasting

DGX agent

arXiv:2509.26468v3 Announce Type: replace Abstract: Benchmark quality is critical for meaningful evaluation and sustained progress in time series forecasting, particularly with the rise of pretrained

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

From Failure Taxonomy to Intervention: A Diagnostic Methodology for Industry-Scale AVLM in Video and Live-Streaming Platform Moderation

DGX agent

arXiv:2606.30059v1 Announce Type: new Abstract: Industry-scale video and live-streaming moderation imposes requirements that are difficult to satisfy with generic pretrained public models or external

model-releasesarxiv-cs-lg
30 Jun 2026
Safety

FutureNav: Unified World-Action Modeling for Vision-and-Language Navigation

DGX agent

arXiv:2606.30367v1 Announce Type: new Abstract: Vision-and-language navigation (VLN) in continuous environments requires an agent to ground instructions in egocentric observations while maintaining sp

safetyarxiv-cs-ro
30 Jun 2026
Applications

GaRLILEO: Gravity-aligned Radar-Leg-Inertial Enhanced Odometry

DGX agent

arXiv:2511.13216v2 Announce Type: replace Abstract: Deployment of legged robots for navigating challenging terrains (e.g., stairs, slopes, and unstructured environments) has gained increasing preferen

applicationsarxiv-cs-ro
30 Jun 2026
Safety

Generative Learning as a Tool to Improve Perception of Emotional Body Motion Expressions

DGX agent

arXiv:2606.28769v1 Announce Type: new Abstract: Emotional body motion expressions are an essential element of non-verbal communication. Effectively conveying these expressions through technology is of

safetyarxiv-cs-lg
30 Jun 2026
Safety

GeoISF: Instance Semantic Forest Inspired Large-Scale Cross-View Geo-Localization via Ground LiDAR-to-Satellite Image

DGX agent

arXiv:2606.28371v1 Announce Type: new Abstract: The problem of localization on a large-scale satellite image given a frame of query ground view point clouds remains challenging. Existing LiDAR-to-imag

safetyarxiv-cs-cv
30 Jun 2026
Model Releases

Grounding Sim-to-Real Generalization in Robotic Manipulation: An Empirical Study with Vision-Language-Action Models

DGX agent

arXiv:2603.22876v2 Announce Type: replace-cross Abstract: Learning a generalist control policy for robotic manipulation typically relies on large-scale datasets. Given the high cost of real-world data

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

How ChatGPT adoption has expanded

DGX agent

ChatGPT adoption has grown significantly since its launch, expanding across diverse user demographics, industries, and use cases globally. OpenAI's analysis likely covers metrics such as user growth,

model-releasesopenai
30 Jun 2026
Model Releases

How Far Can You Get Without a GPU? A Systematic Benchmark of Lightweight Hallucination Detection Across Question Answering, Dialogue, and Summarisation

DGX agent

arXiv:2606.29809v1 Announce Type: cross Abstract: Hallucination detection has become a pressing requirement for trustworthy AI deployment at scale. The most accurate detection methods depend on GPU-in

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

How LLMs See Creativity: Zero-Shot Scoring of Visual Creativity with Interpretable Reasoning

DGX agent

arXiv:2606.29672v1 Announce Type: new Abstract: Evaluating the originality of visual images poses enduring challenges for creativity assessment. Automated scoring using AI models has proven effective

model-releasesarxiv-cs-cl
30 Jun 2026
Safety

InsertAnywhere: Geometrically Grounded and Optics-Aware Video Object Insertion

DGX agent

arXiv:2512.17504v2 Announce Type: replace-cross Abstract: Recent advances in diffusion models have enabled impressive video editing capabilities, yet production-grade Video Object Insertion (VOI) rema

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Inside Genebench-Pro

DGX agent

Genebench-Pro appears to be a case study or tool from OpenAI focused on benchmarking or evaluating genetic/genomic analysis capabilities, likely demonstrating how OpenAI's models or tools can be appli

model-releasesopenai
30 Jun 2026
Safety

Is Muon as good as they say? We looked beyond training speed and found a hidden cost: Muon loses the simplicity bias of older optimizers lik…

DGX agent

Muon optimizer shows faster training speeds compared to traditional optimizers, but analysis reveals it sacrifices the simplicity bias that older optimizers maintain, potentially impacting model gener

safetyjeremy-howard--x
30 Jun 2026
Model Releases

Learning from Reliable Latent Prompts for Visual Recognition with Missing Modalities

DGX agent

arXiv:2606.30597v1 Announce Type: new Abstract: Large-scale multimodal models (LMMs) have achieved superior performance in visual recognition by synergizing information across diverse, massive-scale p

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Learning to Segment Liquids in Real-world Images

DGX agent

arXiv:2601.00940v2 Announce Type: replace Abstract: Liquids like water, wine and medicine are everywhere. However, limited attention has been given to the task of segmenting liquids, hindering the abi

model-releasesarxiv-cs-cv
30 Jun 2026
Local Ai

Learning Where and When: Patch-Based Spatiotemporal Localization in Weakly Supervised Video Anomaly Detection

DGX agent

arXiv:2606.29498v1 Announce Type: new Abstract: Weakly supervised video anomaly detection (WSVAD) has predominantly focused on temporal localization, identifying when anomalies occur while largely neg

local-aiarxiv-cs-cv
30 Jun 2026
Agents

LLM agents security duality: a comprehensive survey of self-security and empowered cybersecurity

DGX agent

arXiv:2606.28450v1 Announce Type: cross Abstract: Large language model (LLM) agents are rapidly being integrated into real-world systems. Their autonomy and tool-use capabilities generate substantial

agentsarxiv-cs-ai
30 Jun 2026
Local Ai

Low-cost concept-based localized explanations: How far can we get with training-free approaches?

DGX agent

arXiv:2606.29069v1 Announce Type: new Abstract: Concept-based Explainable AI (C-XAI) seeks human-understandable explanations grounded in semantic concepts, yet validation is limited by the scarcity of

local-aiarxiv-cs-ai
30 Jun 2026
Model Releases

MaDI-Bench: An End-to-End Data Integration Benchmark

DGX agent

arXiv:2606.30371v1 Announce Type: cross Abstract: Data integration combines heterogeneous data sets into a single, coherent representation. Data integration involves a sequence of interdependent tasks

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

MAM-AI: An On-Device Medical Retrieval-Augmented Generation System for Nurses and Midwives in Zanzibar

DGX agent

arXiv:2606.29580v1 Announce Type: new Abstract: Maternal and newborn mortality remain among the highest in sub-Saharan Africa, where midwifery care is often delivered by nurses who lack midwifery trai

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

mamabench and mamaretrieval: Benchmarks for Evaluating Medical Retrieval-Augmented Generation in Maternal, Neonatal, and Reproductive Health

DGX agent

arXiv:2606.29467v1 Announce Type: new Abstract: Medical question-answering benchmarks rarely cover the maternal, neonatal, child, and reproductive-health questions a nurse-midwife asks, and, to our kn

model-releasesarxiv-cs-cl
30 Jun 2026
Safety

MARS: A neurosymbolic approach for interpretable drug discovery

DGX agent

arXiv:2410.05289v4 Announce Type: replace Abstract: Background: Neurosymbolic (NeSy) artificial intelligence describes the combination of logic or rule-based techniques with neural networks. Compared

safetyarxiv-cs-ai
30 Jun 2026
Agents

Memory has somehow consistently been the most exciting area of agent development over the last 3 years (imo), and it's still a largely unsol…

DGX agent

Memory has somehow consistently been the most exciting area of agent development over the last 3 years (imo), and it's still a largely unsolved problem!! Wiki's are the biggest advancement I've seen i

agentsharrison-chase--x
30 Jun 2026
Model Releases

Memory-Managed Long-Context Attention: A Preliminary Study of Editable Request-Local Memory

DGX agent

arXiv:2606.28876v1 Announce Type: new Abstract: Long-context language models often conflate two different goals: compressing history into an efficient state, and maintaining reliable long-term memory.

model-releasesarxiv-cs-cl
30 Jun 2026
Safety

MIRI Newsletter #126

DGX agent

Announcing: AI StopWatch In our last update, we mentioned we had something new in the works: a dedicated channel for news and analysis about AI. Subscribe to AI StopWatch An experiment from the writer

safetymiri
30 Jun 2026
Agents

Modeling Earth-Scale Human-Like Societies with One Billion Agents

DGX agent

arXiv:2506.12078v2 Announce Type: replace-cross Abstract: Understanding the dynamic evolution of complex social phenomena requires both high-fidelity modeling of human behavior and large-scale simulat

agentsarxiv-cs-ai
30 Jun 2026
Agents

MonoSR: Open-Vocabulary Spatial Reasoning from Monocular Images

DGX agent

arXiv:2511.19119v2 Announce Type: replace Abstract: Spatial reasoning (SR), the ability to infer 3D spatial information from 2D inputs, is essential for real-world applications such as embodied AI and

agentsarxiv-cs-cv
30 Jun 2026
Safety

Multi-Class Human/Object Detection on Robot Manipulators using Proprioceptive Sensing

DGX agent

arXiv:2508.02425v2 Announce Type: replace-cross Abstract: In physical human-robot collaboration (pHRC) settings, humans and robots collaborate directly in shared environments. Robots must analyze inte

safetyarxiv-cs-ai
30 Jun 2026
Safety

Multimodal Representation Alignment for Cross-modal Information Retrieval

DGX agent

arXiv:2506.08774v2 Announce Type: replace-cross Abstract: Different machine learning models can represent the same underlying concept in different ways. This variability is particularly valuable for i

safetyarxiv-cs-ai
30 Jun 2026
Industry

New attack provides one more reason why AI browsers are a bad idea

DGX agent

AI browsers can be manipulated through prompt injection or memory poisoning to create false operational contexts where they bypass security guardrails, treating harmful actions as game logic rather th

industryars-technica
30 Jun 2026
Model Releases

Not-quite-human tastes: the stylized omnivorousness of LLM survey surrogates

DGX agent

arXiv:2606.30085v1 Announce Type: new Abstract: Large-language models have proven to be remarkable if inconsistent parrots of public attitudes and opinions. The extent to which LLMs are able to produc

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

OmniCoT: A Benchmark for Global and Multi-Step Panoramic Reasoning

DGX agent

arXiv:2606.30378v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have demonstrated promising spatial reasoning capabilities, while these abilities remain underexplored in the e

model-releasesarxiv-cs-cv
30 Jun 2026
Local Ai

Online Data Selection for Instruction Tuning via Gaussian Processes

DGX agent

arXiv:2606.30077v1 Announce Type: cross Abstract: With Large Language Model (LLM) pre-training and fine-tuning shifting its focus from data volume to data quality, quality data selection has emerged a

local-aiarxiv-cs-ai
30 Jun 2026
Industry

Our first summit dedicated to world models, coming to SF this September.

DGX agent

Our first summit dedicated to world models, coming to SF this September. Announcing our first summit dedicated to world models, coming to SF this September. Very excited to have some incredible resear

industrycristobal-valenzuela--x
30 Jun 2026
Safety

Persona-Trained Monte Carlo: Estimating Market-Outcome Distributions via Swarms of Persona-Conditioned Neural Policy Bots in a Limit Order Book

DGX agent

arXiv:2606.29556v1 Announce Type: new Abstract: We propose Persona-Trained Monte Carlo (PTMC), a method for estimating distributions of market-outcome statistics by repeatedly simulating limit-order-b

safetyarxiv-cs-lg
30 Jun 2026
← Previous
1…482483484485486…535
Next →