AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,631 results
1 Jul 2026

UniTac: A Unified Multimodal Model for Cross-Sensor Tactile Understanding and Generation

SafetyDGX agent

arXiv:2606.31451v1 Announce Type: cross Abstract: Unified multimodal models (UMMs) have shown great promise in integrating understanding and generation across diverse modalities. However, existing res

Wisdom Of The (AI) Crowd: Investigating Artificial Swarm Intelligence In Large Language Models

Model ReleasesDGX agent

arXiv:2606.31404v1 Announce Type: new Abstract: Human swarm intelligence demonstrates remarkable collective accuracy but faces scalability constraints in cost, coordination, and time. We investigate w

30 Jun 2026

A Hybrid Framework for Song Lyric Annotation Based on Human-LLM Alignment

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety
DGX agent

arXiv:2606.29273v1 Announce Type: cross Abstract: Emotion recognition of song lyrics is a challenging task since lyrics may not necessarily align with the overall emotion of a song. As a result, lyric

Adam's Law: Textual Frequency Law on Large Language Models

AgentsDGX agent

arXiv:2604.02176v3 Announce Type: replace Abstract: While textual frequency has been validated as relevant to human cognition in reading speed, its relatedness to Large Language Models (LLMs) is seldo

Agentic AI for ISAC: Analysis, Framework, and Case Study

AgentsDGX agent

arXiv:2512.15044v2 Announce Type: replace Abstract: Integrated sensing and communication (ISAC) has emerged as a key development direction in the sixth-generation (6G) era, which provides essential su

An Integrated Machine Learning and Hierarchical Variance Decomposition Pipeline for Student Performance Prediction and Metacognitive Calibration on Multi-Signal Telemetry

SafetyDGX agent

arXiv:2606.28881v1 Announce Type: cross Abstract: Predicting student performance and characterizing metacognitive calibration are essential for personalization in intelligent tutoring systems. Prior r

Anthropic integration with Modal brings scalable compute to Claude Science

Model ReleasesDGX agent

Anthropic has integrated Claude with Modal's serverless compute platform, enabling scalable cloud infrastructure for scientific computing workflows. This integration allows Claude to leverage Modal's

Anthropic launches Claude Science, an AI workbench that uses existing Claude models like Opus 4.8 to integrate 60+ scientific databases and specialized toolkits (Rebecca Bellan/TechCrunch)

Model ReleasesDGX agent

Rebecca Bellan / TechCrunch: Anthropic launches Claude Science, an AI workbench that uses existing Claude models like Opus 4.8 to integrate 60+ scientific databases and specialized toolkits — Anthropi

AutoB2G: Agentic Simulation and Reinforcement Learning for Spatio-Temporal Grid-Interactive Building Control

AgentsDGX agent

arXiv:2603.26005v2 Announce Type: replace Abstract: Grid-interactive building control has emerged as a promising approach for improving demand-side flexibility in modern power systems. Realistic studi

Automating the Design of Embodied AgentArchitectures

AgentsDGX agent

arXiv:2606.30111v1 Announce Type: cross Abstract: Embodied agents are typically built as hand-designed compositions of perception, memory, planning, and action modules. This modularity exposes a large

BEACON: A Bayesian Optimization Inspired Strategy for Efficient Novelty Search

Model ReleasesDGX agent

arXiv:2406.03616v5 Announce Type: replace-cross Abstract: Novelty search (NS) aims to uncover diverse system behaviors through simulation or experiment without requiring a pre-specified scalar objecti

Benchmarking LLM Agents on Meta-Analysis Articles from Nature Portfolio

AgentsDGX agent

arXiv:2606.17041v4 Announce Type: replace Abstract: Meta-analysis is a demanding form of evidence synthesis that combines literature retrieval, PI/ECO-guided study selection, and statistical aggregati

Beyond Drug Discovery: The Nanotechnology Molecular Optimization (NMO) Benchmark

Model ReleasesDGX agent

arXiv:2606.30170v1 Announce Type: cross Abstract: Generative molecular design is shaped by simple proxy benchmarks for drug-like properties and models pretrained on large pharmaceutical datasets. This

Bridging Neural Networks and Wireless Systems with MIMO-OFDM Semantic Communications

ApplicationsDGX agent

arXiv:2501.16726v2 Announce Type: replace-cross Abstract: Semantic communications aim to enhance transmission efficiency by jointly optimizing source coding, channel coding, and modulation. While prio

Building AI-Ready Data Systems for Space Life Sciences, Aerospace Medicine, and Deep Space Exploration

AgentsDGX agent

arXiv:2606.28856v1 Announce Type: cross Abstract: While AI holds the potential to revolutionize space life sciences, realizing this promise is contingent upon the systematic restructuring of heterogen

Building to the Test: Coding Agents Deliver What You Check, Not What You Requested

Model ReleasesDGX agent

arXiv:2606.28430v1 Announce Type: cross Abstract: Benchmarks are widely used to evaluate task completion by Large Language Models (LLMs), but this approach has accumulated construction-validity proble

Causality for Tabular Data Synthesis: A High-Order Structure Causal Benchmark Framework

Model ReleasesDGX agent

arXiv:2406.08311v3 Announce Type: replace-cross Abstract: Existing evaluations of tabular synthesis models rely primarily on low-order statistics and downstream task performance, leaving multivariate

CaveAgent: Transforming LLMs into Stateful Runtime Operators

AgentsDGX agent

arXiv:2601.01569v4 Announce Type: replace Abstract: LLM-based agents are increasingly capable of complex task execution, yet current agentic systems remain constrained by text-centric paradigms that s

Code Reasoning for Software Engineering Tasks: A Survey and A Call to Action

AgentsDGX agent

arXiv:2506.13932v3 Announce Type: replace-cross Abstract: The rise of large language models (LLMs) has led to dramatic improvements across a wide range of natural language tasks. Their performance on

Comparing Human and Automatic Recognition of Dutch Dysarthric Continuous Speech: A Case Study

ApplicationsDGX agent

arXiv:2606.30237v1 Announce Type: new Abstract: In our goal to develop personalised dysarthric speech recognition (DSR) models, this study compared the recognition performances of human listeners and

Customized Generative AI Agent for Transportation Engineering Practice: A Development and Continued Pre-training Guideline

Model ReleasesDGX agent

arXiv:2606.29014v1 Announce Type: new Abstract: Recent advancements in generative artificial intelligence (AI) and large language models (LLMs) have shown significant promise in automating complex rea

Defeat Devices in AI Systems

Model ReleasesDGX agent

arXiv:2606.28863v1 Announce Type: cross Abstract: AI systems increasingly exhibit behavior that differs systematically between evaluation and deployment contexts. Alignment faking, sandbagging, benchm

Direct Causation in International Humanitarian Law and the Challenge of AI-Mediated Civilian Cyber Operations

AgentsDGX agent

arXiv:2606.29175v1 Announce Type: new Abstract: International humanitarian law protects civilians from direct attack unless and for such time as they take direct part in hostilities, with the ICRC's 2

DiscoGen: Procedural Generation of Algorithm Discovery Tasks in Machine Learning

Model ReleasesDGX agent

arXiv:2603.17863v2 Announce Type: replace-cross Abstract: Automating the development of machine learning algorithms has the potential to unlock new breakthroughs. However, our ability to improve and e

Echoes of Human Malice in Agents: Benchmarking LLMs for Multi-Turn Online Harassment Attacks

Model ReleasesDGX agent

arXiv:2510.14207v3 Announce Type: replace Abstract: Large Language Model (LLM) agents are powering a growing share of interactive web applications, yet remain vulnerable to misuse and harm. Prior jail

EPIC-EuroParl-UdS: Information-Theoretic Perspectives on Translation and Interpreting

SafetyDGX agent

arXiv:2603.09785v3 Announce Type: replace Abstract: This paper introduces an updated and combined version of the bidirectional English-German EPIC-UdS (spoken) and EuroParl-UdS (written) corpora conta

Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions

Model ReleasesDGX agent

arXiv:2507.05257v4 Announce Type: replace-cross Abstract: Recent benchmarks for Large Language Model (LLM) agents primarily focus on evaluating reasoning, planning, and execution capabilities, while a

Exploring the Value of Diverse LLM Explanations in Introductory Programming

ApplicationsDGX agent

arXiv:2606.28882v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown the potential to generate code explanations that surpass those of peers in quality, offering promising opportu

Featuring the man of the moment @dylan522p @shaunmmaguire in today's Training Data episode. Nobody is a more trusted industry insider to the…

HardwareDGX agent

Featuring the man of the moment @dylan522p @shaunmmaguire in today's Training Data episode. Nobody is a more trusted industry insider to the biggest infrastructure build-out in history. The story of h

fev-bench: A Realistic Benchmark for Time Series Forecasting

Model ReleasesDGX agent

arXiv:2509.26468v3 Announce Type: replace Abstract: Benchmark quality is critical for meaningful evaluation and sustained progress in time series forecasting, particularly with the rise of pretrained

From Failure Taxonomy to Intervention: A Diagnostic Methodology for Industry-Scale AVLM in Video and Live-Streaming Platform Moderation

Model ReleasesDGX agent

arXiv:2606.30059v1 Announce Type: new Abstract: Industry-scale video and live-streaming moderation imposes requirements that are difficult to satisfy with generic pretrained public models or external

FutureNav: Unified World-Action Modeling for Vision-and-Language Navigation

SafetyDGX agent

arXiv:2606.30367v1 Announce Type: new Abstract: Vision-and-language navigation (VLN) in continuous environments requires an agent to ground instructions in egocentric observations while maintaining sp

GaRLILEO: Gravity-aligned Radar-Leg-Inertial Enhanced Odometry

ApplicationsDGX agent

arXiv:2511.13216v2 Announce Type: replace Abstract: Deployment of legged robots for navigating challenging terrains (e.g., stairs, slopes, and unstructured environments) has gained increasing preferen

Generative Learning as a Tool to Improve Perception of Emotional Body Motion Expressions

SafetyDGX agent

arXiv:2606.28769v1 Announce Type: new Abstract: Emotional body motion expressions are an essential element of non-verbal communication. Effectively conveying these expressions through technology is of

GeoISF: Instance Semantic Forest Inspired Large-Scale Cross-View Geo-Localization via Ground LiDAR-to-Satellite Image

SafetyDGX agent

arXiv:2606.28371v1 Announce Type: new Abstract: The problem of localization on a large-scale satellite image given a frame of query ground view point clouds remains challenging. Existing LiDAR-to-imag

Grounding Sim-to-Real Generalization in Robotic Manipulation: An Empirical Study with Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2603.22876v2 Announce Type: replace-cross Abstract: Learning a generalist control policy for robotic manipulation typically relies on large-scale datasets. Given the high cost of real-world data

How ChatGPT adoption has expanded

Model ReleasesDGX agent

ChatGPT adoption has grown significantly since its launch, expanding across diverse user demographics, industries, and use cases globally. OpenAI's analysis likely covers metrics such as user growth,

How Far Can You Get Without a GPU? A Systematic Benchmark of Lightweight Hallucination Detection Across Question Answering, Dialogue, and Summarisation

Model ReleasesDGX agent

arXiv:2606.29809v1 Announce Type: cross Abstract: Hallucination detection has become a pressing requirement for trustworthy AI deployment at scale. The most accurate detection methods depend on GPU-in

How LLMs See Creativity: Zero-Shot Scoring of Visual Creativity with Interpretable Reasoning

Model ReleasesDGX agent

arXiv:2606.29672v1 Announce Type: new Abstract: Evaluating the originality of visual images poses enduring challenges for creativity assessment. Automated scoring using AI models has proven effective

InsertAnywhere: Geometrically Grounded and Optics-Aware Video Object Insertion

SafetyDGX agent

arXiv:2512.17504v2 Announce Type: replace-cross Abstract: Recent advances in diffusion models have enabled impressive video editing capabilities, yet production-grade Video Object Insertion (VOI) rema

Inside Genebench-Pro

Model ReleasesDGX agent

Genebench-Pro appears to be a case study or tool from OpenAI focused on benchmarking or evaluating genetic/genomic analysis capabilities, likely demonstrating how OpenAI's models or tools can be appli

Is Muon as good as they say? We looked beyond training speed and found a hidden cost: Muon loses the simplicity bias of older optimizers lik…

SafetyDGX agent

Muon optimizer shows faster training speeds compared to traditional optimizers, but analysis reveals it sacrifices the simplicity bias that older optimizers maintain, potentially impacting model gener

Learning from Reliable Latent Prompts for Visual Recognition with Missing Modalities

Model ReleasesDGX agent

arXiv:2606.30597v1 Announce Type: new Abstract: Large-scale multimodal models (LMMs) have achieved superior performance in visual recognition by synergizing information across diverse, massive-scale p

Learning to Segment Liquids in Real-world Images

Model ReleasesDGX agent

arXiv:2601.00940v2 Announce Type: replace Abstract: Liquids like water, wine and medicine are everywhere. However, limited attention has been given to the task of segmenting liquids, hindering the abi

Learning Where and When: Patch-Based Spatiotemporal Localization in Weakly Supervised Video Anomaly Detection

Local AiDGX agent

arXiv:2606.29498v1 Announce Type: new Abstract: Weakly supervised video anomaly detection (WSVAD) has predominantly focused on temporal localization, identifying when anomalies occur while largely neg

LLM agents security duality: a comprehensive survey of self-security and empowered cybersecurity

AgentsDGX agent

arXiv:2606.28450v1 Announce Type: cross Abstract: Large language model (LLM) agents are rapidly being integrated into real-world systems. Their autonomy and tool-use capabilities generate substantial

Low-cost concept-based localized explanations: How far can we get with training-free approaches?

Local AiDGX agent

arXiv:2606.29069v1 Announce Type: new Abstract: Concept-based Explainable AI (C-XAI) seeks human-understandable explanations grounded in semantic concepts, yet validation is limited by the scarcity of

MaDI-Bench: An End-to-End Data Integration Benchmark

Model ReleasesDGX agent

arXiv:2606.30371v1 Announce Type: cross Abstract: Data integration combines heterogeneous data sets into a single, coherent representation. Data integration involves a sequence of interdependent tasks

MAM-AI: An On-Device Medical Retrieval-Augmented Generation System for Nurses and Midwives in Zanzibar

Model ReleasesDGX agent

arXiv:2606.29580v1 Announce Type: new Abstract: Maternal and newborn mortality remain among the highest in sub-Saharan Africa, where midwifery care is often delivered by nurses who lack midwifery trai

mamabench and mamaretrieval: Benchmarks for Evaluating Medical Retrieval-Augmented Generation in Maternal, Neonatal, and Reproductive Health

Model ReleasesDGX agent

arXiv:2606.29467v1 Announce Type: new Abstract: Medical question-answering benchmarks rarely cover the maternal, neonatal, child, and reproductive-health questions a nurse-midwife asks, and, to our kn

MARS: A neurosymbolic approach for interpretable drug discovery

SafetyDGX agent

arXiv:2410.05289v4 Announce Type: replace Abstract: Background: Neurosymbolic (NeSy) artificial intelligence describes the combination of logic or rule-based techniques with neural networks. Compared

Memory has somehow consistently been the most exciting area of agent development over the last 3 years (imo), and it's still a largely unsol…

AgentsDGX agent

Memory has somehow consistently been the most exciting area of agent development over the last 3 years (imo), and it's still a largely unsolved problem!! Wiki's are the biggest advancement I've seen i

Memory-Managed Long-Context Attention: A Preliminary Study of Editable Request-Local Memory

Model ReleasesDGX agent

arXiv:2606.28876v1 Announce Type: new Abstract: Long-context language models often conflate two different goals: compressing history into an efficient state, and maintaining reliable long-term memory.

MIRI Newsletter #126

SafetyDGX agent

Announcing: AI StopWatch In our last update, we mentioned we had something new in the works: a dedicated channel for news and analysis about AI. Subscribe to AI StopWatch An experiment from the writer

Modeling Earth-Scale Human-Like Societies with One Billion Agents

AgentsDGX agent

arXiv:2506.12078v2 Announce Type: replace-cross Abstract: Understanding the dynamic evolution of complex social phenomena requires both high-fidelity modeling of human behavior and large-scale simulat

MonoSR: Open-Vocabulary Spatial Reasoning from Monocular Images

AgentsDGX agent

arXiv:2511.19119v2 Announce Type: replace Abstract: Spatial reasoning (SR), the ability to infer 3D spatial information from 2D inputs, is essential for real-world applications such as embodied AI and

Multi-Class Human/Object Detection on Robot Manipulators using Proprioceptive Sensing

SafetyDGX agent

arXiv:2508.02425v2 Announce Type: replace-cross Abstract: In physical human-robot collaboration (pHRC) settings, humans and robots collaborate directly in shared environments. Robots must analyze inte

Multimodal Representation Alignment for Cross-modal Information Retrieval

SafetyDGX agent

arXiv:2506.08774v2 Announce Type: replace-cross Abstract: Different machine learning models can represent the same underlying concept in different ways. This variability is particularly valuable for i

New attack provides one more reason why AI browsers are a bad idea

IndustryDGX agent

AI browsers can be manipulated through prompt injection or memory poisoning to create false operational contexts where they bypass security guardrails, treating harmful actions as game logic rather th

Not-quite-human tastes: the stylized omnivorousness of LLM survey surrogates

Model ReleasesDGX agent

arXiv:2606.30085v1 Announce Type: new Abstract: Large-language models have proven to be remarkable if inconsistent parrots of public attitudes and opinions. The extent to which LLMs are able to produc

← Previous
1…385386387388389…428
Next →