AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,629 results
19 May 2026

Machines Learn Number Fields, But How? The Case of Galois Groups

TutorialsDGX agent

arXiv:2508.06670v2 Announce Type: replace-cross Abstract: By applying interpretable machine learning methods such as decision trees, we study how simple models can classify the Galois groups of Galois

Mitigating Conversational Inertia in Multi-Turn Agents

SafetyDGX agent

arXiv:2602.03664v3 Announce Type: replace Abstract: Large language models excel as few-shot learners when provided with appropriate demonstrations, yet this strength becomes problematic in multiturn a

Modelling Customer Trajectories with Reinforcement Learning for Practical Retail Insights

AgentsDGX agent

arXiv:2605.18449v1 Announce Type: cross Abstract: Understanding customer movement within retail spaces is essential for optimizing store layouts. Real-world trajectory data can provide highly accurate

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Multi-agent AI systems outperform human teams in creativity

Local AiDGX agent

arXiv:2605.17885v1 Announce Type: cross Abstract: Although artificial intelligence (AI) now matches or exceeds human performance across numerous cognitive tasks, creativity remains a highly contested

Multi-Paradigm Agent Interaction in Practice:A Systematic Analysis of Generator-Evaluator, ReAct Loop,and Adversarial Evaluation in the buddyMe Framework

AgentsDGX agent

arXiv:2605.16821v1 Announce Type: new Abstract: The rapid evolution of Large Language Model (LLM) agents has produced diverse interaction paradigms, yet few production systems integrate multiple parad

Multi-site PPG: An In-the-Wild Physiological Dataset from Emerging Multi-site Wearables

Model ReleasesDGX agent

arXiv:2605.17859v1 Announce Type: cross Abstract: Wearables are widely used for mobile health monitoring, and photoplethysmography (PPG) is a key sensing modality for heart rate and related physiologi

Multimodal Cultural Heritage Knowledge Graph Extension with Language and Vision Models

Model ReleasesDGX agent

arXiv:2605.17669v1 Announce Type: new Abstract: The preservation and interpretation of cultural heritage increasingly rely on digital technologies, among which Knowledge Graphs (KGs) stand out for the

New Wide-Net-Casting Jailbreak Attacks Risk Large Models

SafetyDGX agent

arXiv:2605.17128v1 Announce Type: cross Abstract: Jailbreak attacks on large models have drawn growing attention due to their close ties to societal safety. This work identifies a practical yet unexpl

Not What You Asked For: Typographic Attacks in Household Robot Manipulation

Model ReleasesDGX agent

arXiv:2605.18593v1 Announce Type: cross Abstract: Open-vocabulary embodied AI agents increasingly rely on vision-language models such as CLIP for object perception and task grounding. However, the sha

OlmoEarth v1.1: A more efficient family of models

ToolsDGX agent

OlmoEarth v1.1 represents an updated release of Allen AI's open-source language models designed with improved efficiency compared to the original version. The update likely focuses on reduced computat

Ordinal Adaptive Correction: A Data-Centric Approach to Ordinal Image Classification with Noisy Labels

Model ReleasesDGX agent

arXiv:2509.02351v3 Announce Type: replace-cross Abstract: Labeled data is a fundamental component in training supervised deep learning models for computer vision tasks. However, the labeling process,

PAIR: Prefix-Aware Internal Reward Model for Multi-Turn Agent Optimization

SafetyDGX agent

arXiv:2605.17877v1 Announce Type: new Abstract: A significant hurdle for current LLMs is the execution of complex, multi-stage tasks. Group Relative Policy Optimization (GRPO) has been emerging as a l

PersonaArena: Dynamic Simulation for Evaluating and Enhancing Persona-Level Role-Playing in Large Language Models

AgentsDGX agent

arXiv:2605.17044v1 Announce Type: new Abstract: Large language models (LLMs) increasingly serve as interactive social agents, yet their ability to maintain coherent and authentic persona-level role-pl

Politicians who boost AI do not represent the views of their constituents. New polling in the UK shows: - Most people think AI will destroy …

SafetyDGX agent

Politicians who boost AI do not represent the views of their constituents. New polling in the UK shows: - Most people think AI will destroy more jobs than it creates (57% v 17%) - 7 in 10 are worried

Position: AI Evaluations Should be Grounded on a Theory of Capability

Model ReleasesDGX agent

arXiv:2509.19590v2 Announce Type: replace Abstract: Evaluations of generative models are now ubiquitous, and their outcomes critically shape public and scientific expectations of AI's capabilities. Ye

PriHA: A RAG-Enhanced LLM Framework for Primary Healthcare Assistant in Hong Kong

Model ReleasesDGX agent

arXiv:2604.14215v2 Announce Type: replace-cross Abstract: To address the unsustainable rise in public health expenditures, the Hong Kong SAR Government is shifting its strategic focus to primary healt

Progressive Generalization Augmentation with Deeply Coupled RND-PPO and Domain-Prioritized Noise Injection for Robust Crop Management Reinforcement Learning

HardwareDGX agent

arXiv:2605.17428v1 Announce Type: cross Abstract: Our preliminary experiments on gym-DSSAT maize irrigation tasks revealed that +/-2 degrees C temperature noise causes an 11.9% reduction in economic r

Rethinking Code Review in the Age of AI: A Vision for Agentic Code Review

SafetyDGX agent

arXiv:2605.17548v1 Announce Type: cross Abstract: Code review has evolved for decades, from informal peer checking to today's pull request (PR) workflows, yet it remains a largely manual, uneven, and

rsi is here. jesus https://x.com/nickevanjoseph/status/2056760504949842219?s=46

Model ReleasesDGX agent

rsi is here. jesus https://x.com/nickevanjoseph/status/2056760504949842219?s=46 Excited to welcome Andrej to the Pretraining team! He'll be building a team focused on using Claude to accelerate pretra

SAMRI: Segment Any MRI

Local AiDGX agent

arXiv:2510.26635v3 Announce Type: replace-cross Abstract: Summary: SAMRI is an MRI-specialized adaptation of the Segment Anything Model achieving superior whole-body MRI segmentation, particularly for

Sequential Structure in Intraday Futures Data: LSTM vs Gradient Boosting on MNQ

ApplicationsDGX agent

arXiv:2605.17724v1 Announce Type: cross Abstract: This paper compares gradient boosting and long short-term memory (LSTM) architectures for intraday directional prediction in Micro E-Mini Nasdaq 100 f

ShareChat: A Dataset of Chatbot Conversations in the Wild

Model ReleasesDGX agent

arXiv:2512.17843v4 Announce Type: replace-cross Abstract: By evaluating Large Language Models (LLMs) through uniform, text-only interfaces, current academic benchmarks obscure how the unique designs a

SkillGenBench: Benchmarking Skill Generation Pipelines for LLM Agents

Model ReleasesDGX agent

arXiv:2605.18693v1 Announce Type: new Abstract: As LLM agents are increasingly built around reusable skills, a central challenge is no longer only whether agents can use provided skills, but whether t

SocialMemBench: Are AI Memory Systems Ready for Social Group Settings?

Model ReleasesDGX agent

arXiv:2605.17789v1 Announce Type: cross Abstract: Memory systems for AI assistants were built for single-user dialogue and fail characteristically when applied to multi-party social group settings. Th

Sometin Beta Pass Notin (SBPN): Improving Multilingual ASR for Nigerian Languages via Knowledge Distillation

Model ReleasesDGX agent

arXiv:2605.17710v1 Announce Type: new Abstract: Although modern multilingual Automatic Speech Recognition (ASR) systems support several Nigerian languages, their performance consistently lags behind h

SonarSweep: Fusing Sonar and Vision for Robust 3D Reconstruction via Plane Sweeping

ApplicationsDGX agent

arXiv:2511.00392v2 Announce Type: replace-cross Abstract: Accurate 3D reconstruction in visually-degraded underwater environments remains a formidable challenge. Single-modality approaches are insuffi

Supervise Less, See More: Training-free Nuclear Instance Segmentation with Prototype-Guided Prompting

Model ReleasesDGX agent

arXiv:2511.19953v2 Announce Type: replace Abstract: Accurate nuclear instance segmentation is a pivotal task in computational pathology, supporting data-driven clinical insights and facilitating downs

Sustainable Intelligence for the Wild: Democratizing Ecological Monitoring via Knowledge-Adaptive Edge Expert Agents

Local AiDGX agent

arXiv:2605.16671v1 Announce Type: new Abstract: Rapid biodiversity loss underscore the urgency of effective monitoring, yet manual surveys remain resource-intensive. While on-device AI offers a scalab

TailedTS: Benchmark Dataset for Heavy-Tailed Time Series Prediction and Periodicity Quantification

Model ReleasesDGX agent

arXiv:2605.16361v1 Announce Type: cross Abstract: We present TailedTS, a large-scale benchmark dataset derived from Wikipedia hourly page view observations throughout 2024, specifically designed to te

Text2CAD-Bench: A Benchmark for LLM-based Text-to-Parametric CAD Generation

Model ReleasesDGX agent

arXiv:2605.18430v1 Announce Type: new Abstract: Text-to-CAD generation aims to create parametric CAD models from natural language, enabling rapid prototyping and intuitive design workflows. However, e

The Recovery Mechanism: Technology, Education, and What Happens When the Pattern Breaks

ApplicationsDGX agent

arXiv:2605.16283v1 Announce Type: cross Abstract: For centuries, each new technology has automated some layer of cognitive work and been absorbed by education retreating upward to teach the skills mac

Time Series Foundation Models as Strong Baselines in Transportation Forecasting: A Large-Scale Benchmark Analysis

Model ReleasesDGX agent

arXiv:2602.24238v2 Announce Type: replace Abstract: Accurate forecasting of transportation dynamics is essential for urban mobility and infrastructure planning. Although recent work has achieved stron

Toward Robust Multilingual Adaptation of LLMs for Low-Resource Languages

SafetyDGX agent

arXiv:2510.14466v3 Announce Type: replace-cross Abstract: Large language models (LLMs) continue to struggle with low-resource languages, primarily due to limited training data, translation noise, and

Towards Universal Physical Adversarial Attacks via a Joint Multi-Objective and Multi-Model Optimization Framework

SafetyDGX agent

arXiv:2605.17772v1 Announce Type: new Abstract: Physical adversarial attacks often overfit single surrogate models and optimization objectives. While ensemble attacks can mitigate this, existing metho

Universal Time-Series Representation Learning: A Survey

ApplicationsDGX agent

arXiv:2401.03717v4 Announce Type: replace-cross Abstract: Time-series data exists in every corner of real-world systems and services, ranging from satellites in the sky to wearable devices on human bo

18 May 2026

AI Consciousness and Existential Risk

SafetyDGX agent

arXiv:2511.19115v2 Announce Type: replace Abstract: In AI, the existential risk denotes the hypothetical threat posed by an artificial system that would possess both the capability and the objective,

ColPackAgent: Agent-Skill-Guided Hard-Particle Monte Carlo Workflows for Colloidal Packing

Model ReleasesDGX agent

arXiv:2605.15625v1 Announce Type: new Abstract: We introduce ColPackAgent, an agent framework that autonomously runs Monte Carlo simulations of colloidal packing through a Model Context Protocol (MCP)

Context-aware Entity-Relation Extraction for Threat Intelligence Knowledge Graphs

Model ReleasesDGX agent

arXiv:2605.15904v1 Announce Type: new Abstract: Cybersecurity Knowledge Graphs (CKGs) unify diverse Cyber Threat Intelligence (CTI) sources into structured, queryable formats, offering scalable soluti

Decouple Searching from Training: Scaling Data Mixing via Model Merging for Large Language Model Pre-training

Model ReleasesDGX agent

arXiv:2602.00747v2 Announce Type: replace-cross Abstract: Determining an effective data mixture is a key factor in Large Language Model (LLM) pre-training, where models must balance general competence

Detecting Heel Strike and toe off Events Using Kinematic Methods and LSTM Models

SafetyDGX agent

arXiv:2503.00794v2 Announce Type: replace Abstract: Accurate gait event detection is crucial for gait analysis, rehabilitation, and assistive technology, particularly in exoskeleton control, where pre

DexJoCo: A Benchmark and Toolkit for Task-Oriented Dexterous Manipulation on MuJoCo

Model ReleasesDGX agent

arXiv:2605.16257v1 Announce Type: new Abstract: Achieving human-level manipulation requires dexterous robotic hands capable of complex object interactions. Advancing such capabilities further demands

Embracing Biased Transition Matrices for Complementary-Label Learning with Many Classes

SafetyDGX agent

arXiv:2605.15586v1 Announce Type: cross Abstract: Complementary-label learning (CLL) is a weakly supervised paradigm where instances are labeled with classes they do not belong to. Despite a decade of

Every time I ask my 10-year-old to use coding agents, he gets extremely disappointed. It turns out that all he wants is to build his own roc…

AgentsDGX agent

Every time I ask my 10-year-old to use coding agents, he gets extremely disappointed. It turns out that all he wants is to build his own rocket simulator. No amount of context engineering helps. No mo

Fine-Tuning NVIDIA Cosmos Predict 2.5 with LoRA/DoRA for Robot Video Generation

HardwareDGX agent

This guide demonstrates how to fine-tune NVIDIA's Cosmos Predict 2.5 video generation model using parameter-efficient techniques like LoRA (Low-Rank Adaptation) and DoRA (Mixture of Experts-based adap

In his weekly Linux kernel post, Linus Torvalds says 'AI tools are great' but duplicate bug reports have made the security list 'almost entirely unmanageable' (Simon Sharwood/The Register)

IndustryDGX agent

Simon Sharwood / The Register: In his weekly Linux kernel post, Linus Torvalds says “AI tools are great” but duplicate bug reports have made the security list “almost entirely unmanageable” — Multiple

Linked Multi-Model Data on Russian Domestic and Foreign Policy Speeches

SafetyDGX agent

arXiv:2605.15886v1 Announce Type: new Abstract: This paper introduces a dataset of interlinked multimodal political communications from the Russian government, addressing persistent deficiencies in th

Metropolis-Scale Road Network Datasets for Fine-Grained Urban Traffic Modeling

ApplicationsDGX agent

arXiv:2510.02278v2 Announce Type: replace Abstract: Modeling traffic dynamics is a critical challenge for urban computing, with applications from real-time traffic management to infrastructure plannin

MorphoHELM: A Comprehensive Benchmark for Evaluating Representations for Microscopy-Based Morphology Assays

Model ReleasesDGX agent

arXiv:2605.15383v1 Announce Type: new Abstract: Microscopy images contain rich information about how cells respond to perturbations, making them essential to applications like drug screening. To quant

MyoChallenge 2025: A New Benchmark for Human Athletic Intelligence

Model ReleasesDGX agent

arXiv:2605.15650v1 Announce Type: new Abstract: Athletic performance represents the pinnacle of human motor intelligence, demanding rapid choices, precise control, agility, and coordinated physical ex

Optimized Three-Dimensional Photovoltaic Structures with LLM guided Tree Search

AgentsDGX agent

arXiv:2605.16191v1 Announce Type: new Abstract: We present a case study for how AI coding systems can be used to generate novel scientific hypotheses. We combine a generic coding agent (Google's AntiG

parallelcbf: A composable safety-filter and auditability framework for tensor-parallel reinforcement learning

SafetyDGX agent

arXiv:2605.15509v1 Announce Type: new Abstract: While Isaac Lab provides massive parallel UAV simulation, OmniSafe and safe-control-gym provide constrained-RL benchmarks, and CBFKit provides control-b

Quantum Artificial Intelligence for Mission-Critical Systems: Foundations, Architectural Elements, and Future Directions

SafetyDGX agent

arXiv:2511.09884v2 Announce Type: replace Abstract: Mission critical (MC) applications such as defense operations, energy management, cybersecurity, and aerospace control require reliable, determinist

Reinforcement learning for adaptive interior point methods in convex quadratic programming

Model ReleasesDGX agent

arXiv:2509.07404v2 Announce Type: replace-cross Abstract: Quadratic programming is a workhorse of modern nonlinear optimization, control, and data science. Although regularized methods offer convergen

RTL-BenchMT: Dynamic Maintenance of RTL Generation Benchmark Through Agent-Assisted Analysis and Revision

Model ReleasesDGX agent

arXiv:2605.15537v1 Announce Type: new Abstract: This paper introduces RTL-BenchMT, an agentic framework for dynamically maintaining RTL generation benchmarks. Large Language Models (LLMs) assisted aut

SMCEvolve: Principled Scientific Discovery via Sequential Monte Carlo Evolution

TutorialsDGX agent

arXiv:2605.15308v1 Announce Type: new Abstract: LLM-driven program evolution has emerged as a powerful tool for automated scientific discovery, yet existing frameworks offer no principled guide for de

Social-Mamba: Socially-Aware Trajectory Forecasting with State-Space Models

Model ReleasesDGX agent

arXiv:2605.15424v1 Announce Type: new Abstract: Human trajectory forecasting is crucial for safe navigation in crowded environments, requiring models that balance accuracy with computational efficienc

The Open Agent Leaderboard

AgentsDGX agent

The Open Agent Leaderboard is a benchmarking system hosted on Hugging Face that evaluates and ranks AI agents based on their performance across various tasks and capabilities. It provides a standardiz

Ti-iLSTM: A TinyDL Approach for Logic-Level Anomaly Detection in Industrial Water Treatment Systems

Local AiDGX agent

arXiv:2605.15874v1 Announce Type: new Abstract: Industrial Water Treatment Systems (IWTS) are safety critical cyber-physical infrastructures and due to increased connectivity, these systems are expose

TokenButler: Token Importance is Predictable

Model ReleasesDGX agent

arXiv:2503.07518v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) rely on the Key-Value (KV) Cache to store token history, enabling efficient decoding of tokens. As the KV-Cache g

TSBOW -- Traffic Surveillance Benchmark for Occluded Vehicles Under Various Weather Conditions

Model ReleasesDGX agent

arXiv:2602.05414v2 Announce Type: replace Abstract: Global warming has intensified the frequency and severity of extreme weather events, which degrade CCTV signal and video quality while disrupting tr

← Previous
1…403404405406407…428
Next →