AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
90,914Total entries
1Added by human
90,913Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
53,690 results
Model Releases

An Annotation Scheme and Classifier for Personal Facts in Dialogue

DGX agent

arXiv:2605.10339v1 Announce Type: new Abstract: The advancement of Large Language Models (LLMs) has enabled their application in personalized dialogue systems. We present an extended annotation scheme

model-releasesarxiv-cs-cl
12 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

An Empirical Study of Multi-Agent Collaboration for Automated Research

DGX agent

arXiv:2603.29632v2 Announce Type: replace-cross Abstract: As AI agents evolve, the community is rapidly shifting from single Large Language Models (LLMs) to Multi-Agent Systems (MAS) to overcome cogni

model-releasesarxiv-cs-ai
12 May 2026
Tutorials

Annotations Mitigate Post-Training Mode Collapse

DGX agent

arXiv:2605.09995v1 Announce Type: new Abstract: Post-training (via supervised fine-tuning) improves instruction-following, but often induces semantic mode collapse by biasing models toward low-entropy

tutorialsarxiv-cs-cl
12 May 2026
Model Releases

AnyDepth-DETR/-YOLO: Any-depth object detection with a single network

DGX agent

arXiv:2605.09407v1 Announce Type: new Abstract: Modern object detectors are static, fixed-depth networks optimized for a single operating point, requiring separate models for different deployment scen

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

AssayBench: An Assay-Level Virtual Cell Benchmark for LLMs and Agents

DGX agent

arXiv:2605.10876v1 Announce Type: cross Abstract: Recent advances in machine learning and large-scale biological data collections have revived the prospect of building a virtual cell, a computational

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Attention Grounded Enhancement for Visual Document Retrieval

DGX agent

arXiv:2511.13415v2 Announce Type: replace-cross Abstract: Visual document retrieval requires understanding heterogeneous and multi-modal content to satisfy implicit information needs. Recent advances

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

AUHead: Realistic Emotional Talking Head Generation via Action Units Control

DGX agent

arXiv:2602.09534v2 Announce Type: replace Abstract: Realistic talking-head video generation is critical for virtual avatars, film production, and interactive systems. Current methods struggle with nua

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Benchmarking Safety Risks of Knowledge-Intensive Reasoning under Malicious Knowledge Editing

DGX agent

arXiv:2605.10146v1 Announce Type: new Abstract: Large language models (LLMs) increasingly rely on knowledge editing to support knowledge-intensive reasoning, but this flexibility also introduces criti

model-releasesarxiv-cs-ai
12 May 2026
Safety

Break the Brake, Not the Wheel: Untargeted Jailbreak via Entropy Maximization

DGX agent

arXiv:2605.10764v1 Announce Type: cross Abstract: Recent studies show that gradient-based universal image jailbreaks on vision-language models (VLMs) exhibit little or no cross-model transferability,

safetyarxiv-cs-ai
12 May 2026
Model Releases

CADBench: A Multimodal Benchmark for AI-Assisted CAD Program Generation

DGX agent

arXiv:2605.10873v1 Announce Type: cross Abstract: Recovering editable CAD programs from images or 3D observations is central to AI-assisted design, but progress is difficult to measure because existin

model-releasesarxiv-cs-ai
12 May 2026
Safety

CAMAL: Improving Attention Alignment and Faithfulness with Segmentation Masks

DGX agent

arXiv:2605.08325v1 Announce Type: cross Abstract: Many vision datasets now provide segmentation masks in addition to annotated images to support a wide range of tasks. In this work, we propose Class A

safetyarxiv-cs-ai
12 May 2026
Safety

Can LLMs Estimate Student Struggles? Human-AI Difficulty Alignment with Proficiency Simulation for Item Difficulty Prediction

DGX agent

arXiv:2512.18880v2 Announce Type: replace-cross Abstract: Accurate estimation of item (question or task) difficulty is critical for educational assessment but suffers from the cold start problem. Whil

safetyarxiv-cs-ai
12 May 2026
Safety

Can Revealed Preferences Clarify LLM Alignment and Steering?

DGX agent

arXiv:2605.08556v1 Announce Type: new Abstract: LLMs are increasingly used to make or support high-stakes decisions under uncertainty, where alignment depends not only on factual accuracy but on how m

safetyarxiv-cs-lg
12 May 2026
Model Releases

Causal Stories from Sensor Traces: Auditing Epistemic Overreach in LLM-Generated Personal Sensing Explanations

DGX agent

arXiv:2605.08590v1 Announce Type: cross Abstract: LLMs are increasingly used to explain personal sensing data, translating traces of activity and mood into natural-language accounts of why an anomalou

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

CHAINTRIX: A multi-pipeline LLM-augmented framework for automated smart-contract security auditing

DGX agent

arXiv:2605.09350v1 Announce Type: new Abstract: Smart-contract exploits have caused billions of USD in cumulative losses, yet audits remain expensive and slow. Automated tools have emerged to close th

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

ChatbotManip: A Dataset to Facilitate Evaluation and Oversight of Manipulative Chatbot Behaviour

DGX agent

arXiv:2506.12090v2 Announce Type: replace Abstract: This paper introduces ChatbotManip, a novel dataset for studying manipulation in Chatbots. It contains simulated generated conversations between a c

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Classification-Head Bias in Class-Level Machine Unlearning: Diagnosis, Mitigation, and Evaluation

DGX agent

arXiv:2605.08730v1 Announce Type: new Abstract: Class-level machine unlearning aims to remove the influence of specified classes while preserving model utility on retained classes. Existing methods ar

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

CMKL: Modality-Aware Continual Learning for Evolving Biomedical Knowledge Graphs

DGX agent

arXiv:2605.10510v1 Announce Type: cross Abstract: Biomedical knowledge graphs are increasingly large, dynamic, and multimodal, driven by rapid advances in biotechnology such as high-throughput sequenc

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

CodeClinic: Evaluating Automation of Coding Skills for Clinical Reasoning Agents

DGX agent

arXiv:2605.09675v1 Announce Type: new Abstract: Clinical reasoning agents based on large language models (LLMs) aim to automate tasks such as intensive care unit (ICU) monitoring and patient state tra

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Confidence-Guided Diffusion Augmentation for Enhanced Bangla Compound Character Recognition

DGX agent

arXiv:2605.10916v1 Announce Type: cross Abstract: Recognition of handwritten Bangla compound characters remains a challenging problem due to complex character structures, large intra-class variation,

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Context-Augmented Code Generation: How Product Context Improves AI Coding Agent Decision Compliance by 49%

DGX agent

arXiv:2605.08112v1 Announce Type: cross Abstract: AI coding agents powered by large language models can read codebases and produce functional code, but they routinely violate team-specific product dec

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

CrackMeBench: Binary Reverse Engineering for Agents

DGX agent

arXiv:2605.10597v1 Announce Type: cross Abstract: Benchmarks for coding agents increasingly measure source-level software repair, and cybersecurity benchmarks increasingly measure broad capture-the-fl

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

CTQWformer: A CTQW-based Transformer for Graph Classification

DGX agent

arXiv:2605.09486v1 Announce Type: cross Abstract: Graph Neural Networks (GNN) and Transformer-based architectures have achieved remarkable progress in graph learning, yet they still struggle to captur

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

DECO: Sparse Mixture-of-Experts with Dense-Comparable Performance on End-Side Devices

DGX agent

arXiv:2605.10933v1 Announce Type: cross Abstract: While Mixture-of-Experts (MoE) scales model capacity without proportionally increasing computation, its massive total parameter footprint creates sign

model-releasesarxiv-cs-cl
12 May 2026
Tutorials

Deep Arguing

DGX agent

arXiv:2605.10569v1 Announce Type: new Abstract: Deep learning has become the dominant approach for creating high capacity, scalable models across diverse data modalities. However, because these models

tutorialsarxiv-cs-ai
12 May 2026
Model Releases

Deepfake Detection that Generalizes Across Benchmarks

DGX agent

arXiv:2508.06248v4 Announce Type: replace Abstract: The generalization of deepfake detectors to unseen manipulation techniques remains a challenge for practical deployment. Although many approaches ad

model-releasesarxiv-cs-cv
12 May 2026
Applications

Diagnosing and Mitigating Domain Shift in Permission-Based Android Malware Detection

DGX agent

arXiv:2605.09028v1 Announce Type: new Abstract: Machine learning-based Android malware detectors often fail in real-world deployment due to domain shift, where models trained on one data source perfor

applicationsarxiv-cs-lg
12 May 2026
Model Releases

Distributional Spectral Diagnostics for Localizing Grokking Transitions

DGX agent

arXiv:2605.08237v1 Announce Type: new Abstract: In grokking, a model first fits the training data while test accuracy remains low, and only later begins to generalize. We ask whether this transition c

model-releasesarxiv-cs-lg
12 May 2026
Safety

Do Linear Probes Generalize Better in Persona Coordinates?

DGX agent

arXiv:2605.09391v1 Announce Type: new Abstract: It is becoming increasingly necessary to have monitors check for harmful behaviors during language model interactions, but text-only monitoring has not

safetyarxiv-cs-ai
12 May 2026
Applications

Dolphin-CN-Dialect: Where Chinese Dialects Matter

DGX agent

arXiv:2605.08961v1 Announce Type: new Abstract: We present Dolphin-CN-Dialect, a streaming-capable ASR model with a focus on Chinese and dialect-rich scenarios. Compared to the previous version, Dolph

applicationsarxiv-cs-cl
12 May 2026
Model Releases

Done, But Not Sure: Disentangling World Completion from Self-Termination in Embodied Agents

DGX agent

arXiv:2605.08747v1 Announce Type: new Abstract: Standard embodied evaluations do not independently score whether an agent correctly commits to task completion at episode closure, a capacity we call te

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Don't Click That: Teaching Web Agents to Resist Deceptive Interfaces

DGX agent

arXiv:2605.09497v1 Announce Type: new Abstract: Vision-language model (VLM) based web agents demonstrate impressive autonomous GUI interaction but remain vulnerable to deceptive interface elements. Ex

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Edge-specific signal propagation on mature chromophore-region 3D mechanism graphs for fluorescent protein quantum-yield prediction

DGX agent

arXiv:2605.06644v2 Announce Type: replace Abstract: Fluorescent protein quantum yield (QY) is governed by the mature chromophore and its three-dimensional microenvironment rather than sequence identit

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Efficient Ensemble Selection from Binary and Pairwise Feedback

DGX agent

arXiv:2605.09588v1 Announce Type: cross Abstract: Organizations increasingly deploy multiple AI systems across task domains, but selecting a small, high-performing ensemble can require costly model ca

model-releasesarxiv-cs-ai
12 May 2026
Applications

Elastic MoE: Unlocking the Inference-Time Scalability of Mixture-of-Experts

DGX agent

arXiv:2509.21892v2 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) models typically fix the number of activated experts k at both training and inference. However, real-world deployment

applicationsarxiv-cs-ai
12 May 2026
Research

ELF: Embedded Language Flows

DGX agent

arXiv:2605.10938v1 Announce Type: cross Abstract: Diffusion and flow-based models have become the de facto approaches for generating continuous data, e.g., in domains such as images and videos. Their

researcharxiv-cs-ai
12 May 2026
Model Releases

ER-Reason: A Benchmark Dataset for LLM Clinical Reasoning in the Emergency Room

DGX agent

arXiv:2505.22919v3 Announce Type: replace Abstract: Existing benchmarks for evaluating the clinical reasoning capabilities of large language models (LLMs) often lack a clear definition of 'clinical re

model-releasesarxiv-cs-cl
12 May 2026
Safety

ETS: Energy-Guided Test-Time Scaling for Training-Free RL Alignment

DGX agent

arXiv:2601.21484v2 Announce Type: replace Abstract: Reinforcement Learning (RL) post-training alignment for language models is effective, but also costly and unstable in practice, owing to its complic

safetyarxiv-cs-lg
12 May 2026
Research

Factual recall in linear associative memories: sharp asymptotics and mechanistic insights

DGX agent

arXiv:2605.10795v1 Announce Type: cross Abstract: Large language models demonstrate remarkable ability in factual recall, yet the fundamental limits of storing and retrieving input--output association

researcharxiv-cs-lg
12 May 2026
Model Releases

Follow the Mean: Reference-Guided Flow Matching

DGX agent

arXiv:2605.10302v1 Announce Type: new Abstract: Existing approaches to controllable generation typically rely on fine-tuning, auxiliary networks, or test-time search. We show that flow matching admits

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Forge: Quality-Aware Reinforcement Learning for NP-Hard Optimization in LLMs

DGX agent

arXiv:2605.08905v1 Announce Type: new Abstract: Large Language Models (LLMs) have achieved remarkable success on reasoning benchmarks through Reinforcement Learning with Verifiable Rewards (RLVR), exc

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Frame In, Frame Out: Measuring Framing Bias in LLM-Generated News Summaries

DGX agent

arXiv:2505.05406v2 Announce Type: replace Abstract: News headlines and summaries shape how events are interpreted through selective emphasis and omission, a phenomenon commonly referred to as framing.

model-releasesarxiv-cs-cl
12 May 2026
Safety

Frequency Adapter with SAM for Generalized Medical Image Segmentation

DGX agent

arXiv:2605.09925v1 Announce Type: new Abstract: Medical image segmentation is a critical task in computer-aided diagnosis and treatment planning. However, deep learning models often struggle to genera

safetyarxiv-cs-cv
12 May 2026
Model Releases

Generating Symmetric Materials using Latent Flow Matching

DGX agent

arXiv:2605.10115v1 Announce Type: new Abstract: Tackling the task of materials generation, we aim to enhance the previously proposed All-atom Diffusion Transformer (ADiT) by introducing SymADiT, a sym

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

GenMed: A Pairwise Generative Reformulation of Medical Diagnostic Tasks

DGX agent

arXiv:2605.10645v1 Announce Type: new Abstract: Data-driven medical AI is traditionally formulated as a discriminative mapping from input X to output Y via a learned function f, which does not general

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

GONE: Structural Knowledge Unlearning via Neighborhood-Expanded Distribution Shaping

DGX agent

arXiv:2603.12275v1 Announce Type: cross Abstract: Unlearning knowledge is a pressing and challenging task in Large Language Models (LLMs) because of their unprecedented capability to memorize and dige

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

GravityGraphSAGE: Link Prediction in Directed Attributed Graphs

DGX agent

arXiv:2605.09408v1 Announce Type: new Abstract: Link prediction (inferring missing or future connections between nodes in a graph) is a fundamental problem in network science with widespread applicati

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

HiDrive: A Closed-Loop Benchmark for High-Level Autonomous Driving

DGX agent

arXiv:2605.09972v1 Announce Type: cross Abstract: End-to-end autonomous driving has witnessed rapid progress, yet existing benchmarks are increasingly saturated, with state-of-the-art models achieving

model-releasesarxiv-cs-cv
12 May 2026
← Previous
1…448449450451452…1119
Next →