AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,356
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,165
  • Local Ai4,929
  • Model Releases23,869
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,356
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,165
  • Local Ai4,929
  • Model Releases23,869
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,356Total entries
1Added by human
88,355Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,584 results
21 Apr 2026

What's Left Unsaid? Detecting and Correcting Misleading Omissions in Multimodal News Previews

Model ReleasesDGX agent

arXiv:2601.05563v2 Announce Type: replace Abstract: Even when factually correct, social-media news previews (image-headline pairs) can induce interpretation drift: by selectively omitting crucial cont

When W4A4 Breaks Camouflaged Object Detection: Token-Group Dual-Constraint Activation Quantization

Model ReleasesDGX agent

arXiv:2604.16855v1 Announce Type: new Abstract: Camouflaged object detection (COD) segments objects that intentionally blend with the background, so predictions depend on subtle texture and boundary c

Where to Focus: Query-Modulated Multimodal Keyframe Selection for Long Video Understanding

Local AiDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.17422v1 Announce Type: new Abstract: Long video understanding remains a formidable challenge for Multimodal Large Language Models (MLLMs) due to the prohibitive computational cost of proces

Why Agents Compromise Safety Under Pressure

SafetyDGX agent

arXiv:2603.14975v2 Announce Type: replace-cross Abstract: Large Language Model agents deployed in complex environments frequently encounter a conflict between maximizing goal achievement and adhering

XOXO: Stealthy Cross-Origin Context Poisoning Attacks against AI Coding Assistants

Model ReleasesDGX agent

arXiv:2503.14281v4 Announce Type: replace-cross Abstract: AI coding assistants are widely used for tasks like code generation. These tools now require large and complex contexts, automatically sourced

XRePIT: A deep learning-computational fluid dynamics hybrid framework implemented in OpenFOAM for fast, robust, and scalable unsteady simulations

Model ReleasesDGX agent

arXiv:2510.21804v2 Announce Type: replace Abstract: Autoregressive neural surrogates offer computational acceleration for fluid dynamics but inherently suffer from error accumulation and non-physical

20 Apr 2026

Accelerate Generative AI Inference on Amazon SageMaker AI with G7e Instances

Model ReleasesDGX agent

Today, we are thrilled to announce the availability of G7e instances powered by NVIDIA RTX PRO 6000 Blackwell Server Edition GPUs on Amazon SageMaker AI. You can provision nodes with 1, 2, 4, and 8 RT

ARC-AGI-3: A New Challenge for Frontier Agentic Intelligence

Model ReleasesDGX agent

arXiv:2603.24621v2 Announce Type: replace Abstract: We introduce ARC-AGI-3, an interactive benchmark for studying agentic intelligence through novel, abstract, turn-based environments in which agents

AST: Adaptive, Seamless, and Training-Free Precise Speech Editing

ResearchDGX agent

arXiv:2604.16056v1 Announce Type: cross Abstract: Text-based speech editing aims to modify specific segments while preserving speaker identity and acoustic context. Existing methods rely on task-speci

Beyond Text Prompts: Precise Concept Erasure through Text-Image Collaboration

ResearchDGX agent

arXiv:2604.15829v1 Announce Type: new Abstract: Text-to-image generative models have achieved impressive fidelity and diversity, but can inadvertently produce unsafe or undesirable content due to impl

Bridging the phenotype-target gap for molecular generation via multi-objective reinforcement learning

TutorialsDGX agent

arXiv:2509.21010v2 Announce Type: replace-cross Abstract: The de novo generation of drug-like molecules capable of inducing desirable phenotypic changes is receiving increasing attention. However, pre

C-Mining: Unsupervised Discovery of Seeds for Cultural Data Synthesis via Geometric Misalignment

SafetyDGX agent

arXiv:2604.15675v1 Announce Type: new Abstract: Achieving cultural alignment in Large Language Models (LLMs) increasingly depends on synthetic data generation. For such synthesis, the most vital initi

Chain-of-Thought Degrades Visual Spatial Reasoning Capabilities of Multimodal LLMs

ResearchDGX agent

arXiv:2604.16060v1 Announce Type: cross Abstract: Multimodal Reasoning Models (MRMs) leveraging Chain-of-Thought (CoT) based thinking have revolutionized mathematical and logical problem-solving. Howe

ChemGraph-XANES: An Agentic Framework for XANES Simulation and Analysis

Model ReleasesDGX agent

arXiv:2604.16205v1 Announce Type: cross Abstract: Computational X-ray absorption near-edge structure (XANES) is widely used to probe local coordination environments, oxidation states, and electronic s

Classic study gave 146 economist teams the same dataset & got wildly different answers New paper reruns it with agentic AI. Claude Code & Co…

Model ReleasesDGX agent

Classic study gave 146 economist teams the same dataset & got wildly different answers New paper reruns it with agentic AI. Claude Code & Codex land near the human median, but with far tighter dispers

Collaborative Filtering Through Weighted Similarities of User and Item Embeddings

ResearchDGX agent

arXiv:2604.15573v1 Announce Type: cross Abstract: In recent years, neural networks and other complex models have dominated recommender systems, often setting new benchmarks for state-of-the-art perfor

Comparing the latent features of universal machine-learning interatomic potentials

SafetyDGX agent

arXiv:2512.05717v3 Announce Type: replace-cross Abstract: The past few years have seen the development of ``universal'' machine-learning interatomic potentials (uMLIPs) capable of approximating the gr

ConlangCrafter: Constructing Languages with a Multi-Hop LLM Pipeline

ResearchDGX agent

arXiv:2508.06094v4 Announce Type: replace Abstract: Constructed languages (conlangs) such as Esperanto and Quenya have played diverse roles in art, philosophy, and international communication. Meanwhi

CSLE: A Reinforcement Learning Platform for Autonomous Security Management

AgentsDGX agent

arXiv:2604.15590v1 Announce Type: cross Abstract: Reinforcement learning is a promising approach to autonomous and adaptive security management in networked systems. However, current reinforcement lea

Cut Your Losses! Learning to Prune Paths Early for Efficient Parallel Reasoning

ApplicationsDGX agent

arXiv:2604.16029v1 Announce Type: new Abstract: Parallel reasoning enhances Large Reasoning Models (LRMs) but incurs prohibitive costs due to futile paths caused by early errors. To mitigate this, pat

CXR-LT 2026 Challenge: Multi-Center Long-Tailed and Zero Shot Chest X-ray Classification

Model ReleasesDGX agent

arXiv:2604.15555v1 Announce Type: new Abstract: Chest X-ray (CXR) interpretation is hindered by the long-tailed distribution of pathologies and the open-world nature of clinical environments. Existing

DB-FGA-Net: Dual Backbone Frequency Gated Attention Network for Multi-Class Brain Tumor Classification with Grad-CAM Interpretability

Local AiDGX agent

arXiv:2510.20299v3 Announce Type: replace-cross Abstract: Brain tumors are a challenging problem in neuro-oncology, where early and precise diagnosis is important for successful treatment. Deep learni

DepCap: Adaptive Block-Wise Parallel Decoding for Efficient Diffusion LM Inference

Local AiDGX agent

arXiv:2604.15750v1 Announce Type: cross Abstract: Diffusion language models (DLMs) have emerged as a promising alternative to autoregressive language generation due to their potential for parallel dec

Enabling Predictive Maintenance in District Heating Substations: A Labelled Dataset and Fault Detection Evaluation Framework based on Service Data

Model ReleasesDGX agent

arXiv:2511.14791v2 Announce Type: replace-cross Abstract: Early detection of faults in district heating substations is imperative to reduce return temperatures and enhance efficiency. However, progres

EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis

AgentsDGX agent

arXiv:2601.05808v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are expected to be trained to act as agents in various real-world environments, but this process relies on rich a

EvoTest: Evolutionary Test-Time Learning for Self-Improving Agentic Systems

Model ReleasesDGX agent

arXiv:2510.13220v2 Announce Type: replace Abstract: A fundamental limitation of current AI agents is their inability to learn complex skills on the fly at test time, often behaving like 'clever but cl

FL-MHSM: Spatially-adaptive Fusion and Ensemble Learning for Flood-Landslide Multi-Hazard Susceptibility Mapping at Regional Scale

ResearchDGX agent

arXiv:2604.16265v1 Announce Type: new Abstract: Existing multi-hazard susceptibility mapping (MHSM) studies often rely on spatially uniform models, treat hazards independently, and provide limited rep

GAViD: A Large-Scale Multimodal Dataset for Context-Aware Group Affect Recognition from Videos

ApplicationsDGX agent

arXiv:2604.16214v1 Announce Type: new Abstract: Understanding affective dynamics in real-world social systems is fundamental to modeling and analyzing human-human interactions in complex environments.

GenHSI: Controllable Generation of Human-Scene Interaction Videos

ResearchDGX agent

arXiv:2506.19840v2 Announce Type: replace Abstract: Large-scale pre-trained video diffusion models have exhibited remarkable capabilities in diverse video generation. However, existing solutions face

GroupDPO: Memory efficient Group-wise Direct Preference Optimization

SafetyDGX agent

arXiv:2604.15602v1 Announce Type: new Abstract: Preference optimization is widely used to align Large Language Models (LLMs) with preference feedback. However, most existing methods train on a single

Hallucination as Trajectory Commitment: Causal Evidence for Asymmetric Attractor Dynamics in Transformer Generation

ResearchDGX agent

arXiv:2604.15400v1 Announce Type: cross Abstract: We present causal evidence that hallucination in autoregressive language models is an early trajectory commitment governed by asymmetric attractor dyn

Kimi K2.6 just entered the chat. > beats gpt 5.4 and opus 4.6 on coding. > open source. open weight. > 3x-5x cheaper. > really good for peop…

AgentsDGX agent

Kimi K2.6 just entered the chat. > beats gpt 5.4 and opus 4.6 on coding. > open source. open weight. > 3x-5x cheaper. > really good for people running agents. > insane at running long tasks. > we're t

@Kimi_Moonshot Kimi-K2.6 just dropped in Qoder. SOTA coding · Long-horizon execution · Agent swarms At 0.3x credits.

AgentsDGX agent

Kimi-K2.6 is a new AI model released through Qoder that achieves state-of-the-art performance in coding tasks and supports long-horizon execution and agent swarms, available at a reduced cost of 0.3x

Learning Behaviorally Grounded Item Embeddings via Personalized Temporal Contexts

ResearchDGX agent

arXiv:2604.15581v1 Announce Type: cross Abstract: Effective user modeling requires distinguishing between short-term and long-term preference evolution. While item embeddings have become a key compone

Learning to Reason with Insight for Informal Theorem Proving

ResearchDGX agent

arXiv:2604.16278v1 Announce Type: new Abstract: Although most of the automated theorem-proving approaches depend on formal proof systems, informal theorem proving can align better with large language

Majority Voting for Code Generation

ResearchDGX agent

arXiv:2604.15618v1 Announce Type: new Abstract: We investigate Functional Majority Voting (FMV), a method based on functional consensus for code generation with Large Language Models, which identifies

MMGait: Towards Multi-Modal Gait Recognition

Model ReleasesDGX agent

arXiv:2604.15979v1 Announce Type: new Abstract: Gait recognition has emerged as a powerful biometric technique for identifying individuals at a distance without requiring user cooperation. Most existi

OpenAI rolls out Chronicle, which builds memories from screen captures to make Codex more aware of context, as a research preview for Pro subscribers on macOS (Zac Hall/9to5Mac)

Model ReleasesDGX agent

Zac Hall / 9to5Mac: OpenAI rolls out Chronicle, which builds memories from screen captures to make Codex more aware of context, as a research preview for Pro subscribers on macOS — Last week, OpenAI r

opus 4.7 seems to have a much better time in claude code if you run without most of the system prompt (claude --system-prompt '.')

Model ReleasesDGX agent

A user reports that Opus 4.7 performs better in Claude Code when executed with a minimal system prompt (using just a period) rather than the full default system prompt, suggesting that reducing system

OT on the Map: Quantifying Domain Shifts in Geographic Space

TutorialsDGX agent

arXiv:2604.16220v1 Announce Type: new Abstract: In computer vision and machine learning for geographic data, out-of-domain generalization is a pervasive challenge, arising from uneven global data cove

PINNACLE: An Open-Source Computational Framework for Classical and Quantum PINNs

Model ReleasesDGX agent

arXiv:2604.15645v1 Announce Type: new Abstract: We present PINNACLE, an open-source computational framework for physics-informed neural networks (PINNs) that integrates modern training strategies, mul

Placing Puzzle Pieces Where They Matter: A Question Augmentation Framework for Reinforcement Learning

ResearchDGX agent

arXiv:2604.15830v1 Announce Type: new Abstract: Reinforcement learning has become a powerful approach for enhancing large language model reasoning, but faces a fundamental dilemma: training on easy pr

Polyglot: Multilingual Style Preserving Speech-Driven Facial Animation

ApplicationsDGX agent

arXiv:2604.16108v1 Announce Type: new Abstract: Speech-Driven Facial Animation (SDFA) has gained significant attention due to its applications in movies, video games, and virtual reality. However, mos

PRL-Bench: A Comprehensive Benchmark Evaluating LLMs' Capabilities in Frontier Physics Research

Model ReleasesDGX agent

arXiv:2604.15411v1 Announce Type: cross Abstract: The paradigm of agentic science requires AI systems to conduct robust reasoning and engage in long-horizon, autonomous exploration. However, current s

Prompt-Driven Code Summarization: A Systematic Literature Review

Local AiDGX agent

arXiv:2604.15385v1 Announce Type: cross Abstract: Software documentation is essential for program comprehension, developer onboarding, code review, and long-term maintenance. Yet producing quality doc

Reading Between the Lines: The One-Sided Conversation Problem

ApplicationsDGX agent

arXiv:2511.03056v2 Announce Type: replace-cross Abstract: Conversational AI is constrained in many real-world settings where only one side of a dialogue can be recorded, such as telemedicine, call cen

Reasoning over Video: Evaluating How MLLMs Extract, Integrate, and Reconstruct Spatiotemporal Evidence

Model ReleasesDGX agent

arXiv:2603.13091v2 Announce Type: replace Abstract: The growing interest in embodied agents increases the demand for spatiotemporal video understanding, yet existing benchmarks largely emphasize extra

Reversible Residual Normalization Alleviates Spatio-Temporal Distribution Shift

TutorialsDGX agent

arXiv:2604.15838v1 Announce Type: new Abstract: Distribution shift severely degrades the performance of deep forecasting models. While this issue is well-studied for individual time series, it remains

Scalable Unseen Objects 6-DoF Absolute Pose Estimation with Robotic Integration

SafetyDGX agent

arXiv:2503.05578v4 Announce Type: replace Abstract: Pose estimation-guided unseen object 6-DoF robotic manipulation is a key task in robotics. However, the scalability of current pose estimation metho

SENSE: Stereo OpEN Vocabulary SEmantic Segmentation

AgentsDGX agent

arXiv:2604.15946v1 Announce Type: new Abstract: Open-vocabulary semantic segmentation enables models to segment objects or image regions beyond fixed class sets, offering flexibility in dynamic enviro

Sentiment Analysis of German Sign Language Fairy Tales

ResearchDGX agent

arXiv:2604.16138v1 Announce Type: new Abstract: We present a dataset and a model for sentiment analysis of German sign language (DGS) fairy tales. First, we perform sentiment analysis for three levels

Sheer fabric with light transmission. Cloth physics that responds to wind. Depth-of-field compositing. Physically-based lighting - all rende…

Model ReleasesDGX agent

This post discusses advanced rendering techniques for realistic visual effects, including sheer fabric simulation with light transmission properties, cloth physics responsive to wind forces, depth-of-

Stein Variational Black-Box Combinatorial Optimization

Model ReleasesDGX agent

arXiv:2604.15837v1 Announce Type: new Abstract: Combinatorial black-box optimization in high-dimensional settings demands a careful trade-off between exploiting promising regions of the search space a

Subliminal Transfer of Unsafe Behaviors in AI Agent Distillation

SafetyDGX agent

arXiv:2604.15559v1 Announce Type: new Abstract: Recent work on subliminal learning demonstrates that language models can transmit semantic traits through data that is semantically unrelated to those t

The Crutch or the Ceiling? How Different Generations of LLMs Shape EFL Student Writings

ApplicationsDGX agent

arXiv:2604.15460v1 Announce Type: cross Abstract: The rapid evolution of Large Language Models (LLMs) has made them powerful tools for enhancing student writing. This study explores the extent and lim

The Metacognitive Monitoring Battery: A Cross-Domain Benchmark for LLM Self-Monitoring

Model ReleasesDGX agent

arXiv:2604.15702v1 Announce Type: new Abstract: We introduce a cross-domain behavioural assay of monitoring-control coupling in LLMs, grounded in the Nelson and Narens (1990) metacognitive framework a

To LLM, or Not to LLM: How Designers and Developers Navigate LLMs as Tools or Teammates

ResearchDGX agent

arXiv:2604.15344v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly integrated into design and development workflows, yet decisions about their use are rarely binary or pur

Towards Robust Endogenous Reasoning: Unifying Drift Adaptation in Non-Stationary Tuning

SafetyDGX agent

arXiv:2604.15705v1 Announce Type: new Abstract: Reinforcement Fine-Tuning (RFT) has established itself as a critical paradigm for the alignment of Multi-modal Large Language Models (MLLMs) with comple

🔬 Training Transformers to solve 95% failure rate of Cancer Trials — Ron Alfa & Daniel Bear, Noetik

ToolsDGX agent

Noetik uses transformer neural networks and machine learning to improve cancer drug trial success rates, addressing the historically high failure rate (95%) in clinical development. The approach lever

Trajectory Planning for Safe Dual Control with Active Exploration

SafetyDGX agent

arXiv:2604.15507v1 Announce Type: new Abstract: Planning safe trajectories under model uncertainty is a fundamental challenge. Robust planning ensures safety by considering worst-case realizations, ye

← Previous
1…579580581582583…1060
Next →