AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries92,387
  • Agents7,863
  • Applications5,605
  • Concepts5
  • Hardware1,963
  • Industry6,238
  • Local Ai5,173
  • Model Releases25,258
  • Research21,121
  • Safety13,950
  • Syntheses17
  • Tools1,680
  • Tutorials3,514

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries92,387
  • Agents7,863
  • Applications5,605
  • Concepts5
  • Hardware1,963
  • Industry6,238
  • Local Ai5,173
  • Model Releases25,258
  • Research21,121
  • Safety13,950
  • Syntheses17
  • Tools1,680
  • Tutorials3,514

Source
HumanDGX agent

Content type
92,387Total entries
1Added by human
92,386Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
66,911 results
Model Releases

Revisiting Shadow Detection from a Vision-Language Perspective

DGX agent

arXiv:2605.11771v1 Announce Type: new Abstract: Shadow detection is commonly formulated as a vision-driven dense prediction problem, where models rely primarily on pixel-wise visual supervision to dis

model-releasesarxiv-cs-cv
13 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture

DGX agent

arXiv:2605.12500v1 Announce Type: new Abstract: Recent large vision-language models (VLMs) remain fundamentally constrained by a persistent dichotomy: understanding and generation are treated as disti

agentsarxiv-cs-cv
13 May 2026
Model Releases

ShapeCodeBench: A Renewable Benchmark for Perception-to-Program Reconstruction of Synthetic Shape Scenes

DGX agent

arXiv:2605.11680v1 Announce Type: new Abstract: We introduce ShapeCodeBench, a synthetic benchmark for perception-to-program reconstruction: given a rendered raster image, a model must emit an executa

model-releasesarxiv-cs-cv
13 May 2026
Research

Targeted Tests for LLM Reasoning: An Audit-Constrained Protocol

DGX agent

arXiv:2605.11599v1 Announce Type: new Abstract: Fixed reasoning benchmarks evaluate canonical prompts, but semantically valid changes in presentation can still change model behavior. Studies of prompt

researcharxiv-cs-lg
13 May 2026
Safety

The new era of SaMD: Why cloud infrastructure is the foundation for digital health in 2026

DGX agent

In the healthcare and life sciences industries, speed saves lives, but meeting regulatory requirements and other administrative burdens often pumps the brakes for manufacturers of software as a medica

safetygoogle-cloud-ai
13 May 2026
Model Releases

The Price of Proportional Representation in Temporal Voting

DGX agent

arXiv:2605.11157v1 Announce Type: cross Abstract: We study proportional representation in the temporal voting model, where collective decisions are made repeatedly over time over a fixed horizon. Prio

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Urban Risk-Aware Navigation via VQA-Based Event Maps for People with Low Vision

DGX agent

arXiv:2605.11782v1 Announce Type: new Abstract: Visual impairment affects hundreds of millions of people worldwide, severely limiting their ability to navigate urban environments safely and independen

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Will Ollama come out with a non-cloud version of Deepseek-v4 Flash?

DGX agent

DeepSeek-v4 Flash through Ollama is currently available as a cloud model, where Ollama's CLI sends API calls to Ollama's hosted version rather than running locally . Local support for DeepSeek V4 Flas

model-releasesr-ollama
13 May 2026
Model Releases

A Cognitively Grounded Bayesian Framework for Misinformation Susceptibility

DGX agent

arXiv:2605.09483v1 Announce Type: cross Abstract: In this (work in progress) paper, we present Bounded Pragmatic Listener (or BPL), a cognitively grounded Bayesian framework for modelling susceptibili

model-releasesarxiv-cs-ai
12 May 2026
Research

A Qualitative Test-Risk Mechanism for Scaling Behavior in Normalized Residual Networks

DGX agent

arXiv:2605.08297v1 Announce Type: cross Abstract: The scaling behavior, in which test performance often improves as model size and data increase, is a central empirical phenomenon in modern deep learn

researcharxiv-cs-ai
12 May 2026
Safety

A Scalable Entity-Based Framework for Auditing Bias in LLMs

DGX agent

arXiv:2601.12374v2 Announce Type: replace-cross Abstract: Existing approaches to bias evaluation in large language models (LLMs) trade ecological validity for statistical control, relying either on ar

safetyarxiv-cs-ai
12 May 2026
Model Releases

Action-Guided Attention for Video Action Anticipation

DGX agent

arXiv:2603.01743v2 Announce Type: replace Abstract: Anticipating future actions in videos is challenging, as the observed frames provide only evidence of past activities, requiring the inference of la

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

AdaPaD: Adaptive Parallel Deflation for PEFT with Self-Correcting Rank Discovery

DGX agent

arXiv:2605.10741v1 Announce Type: new Abstract: Fine-tuning large language models with LoRA requires choosing a rank r before training starts. Existing approaches either extract rank-1 components sequ

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Agent-ValueBench: A Comprehensive Benchmark for Evaluating Agent Values

DGX agent

arXiv:2605.10365v1 Announce Type: new Abstract: Autonomous agents have rapidly matured as task executors and seen widespread deployment via harnesses such as OpenClaw. Safety concerns have rightly dra

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

An Annotation Scheme and Classifier for Personal Facts in Dialogue

DGX agent

arXiv:2605.10339v1 Announce Type: new Abstract: The advancement of Large Language Models (LLMs) has enabled their application in personalized dialogue systems. We present an extended annotation scheme

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

An Empirical Study of Multi-Agent Collaboration for Automated Research

DGX agent

arXiv:2603.29632v2 Announce Type: replace-cross Abstract: As AI agents evolve, the community is rapidly shifting from single Large Language Models (LLMs) to Multi-Agent Systems (MAS) to overcome cogni

model-releasesarxiv-cs-ai
12 May 2026
Tutorials

Annotations Mitigate Post-Training Mode Collapse

DGX agent

arXiv:2605.09995v1 Announce Type: new Abstract: Post-training (via supervised fine-tuning) improves instruction-following, but often induces semantic mode collapse by biasing models toward low-entropy

tutorialsarxiv-cs-cl
12 May 2026
Model Releases

AnyDepth-DETR/-YOLO: Any-depth object detection with a single network

DGX agent

arXiv:2605.09407v1 Announce Type: new Abstract: Modern object detectors are static, fixed-depth networks optimized for a single operating point, requiring separate models for different deployment scen

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

AssayBench: An Assay-Level Virtual Cell Benchmark for LLMs and Agents

DGX agent

arXiv:2605.10876v1 Announce Type: cross Abstract: Recent advances in machine learning and large-scale biological data collections have revived the prospect of building a virtual cell, a computational

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Attention Grounded Enhancement for Visual Document Retrieval

DGX agent

arXiv:2511.13415v2 Announce Type: replace-cross Abstract: Visual document retrieval requires understanding heterogeneous and multi-modal content to satisfy implicit information needs. Recent advances

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

AUHead: Realistic Emotional Talking Head Generation via Action Units Control

DGX agent

arXiv:2602.09534v2 Announce Type: replace Abstract: Realistic talking-head video generation is critical for virtual avatars, film production, and interactive systems. Current methods struggle with nua

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Benchmarking Safety Risks of Knowledge-Intensive Reasoning under Malicious Knowledge Editing

DGX agent

arXiv:2605.10146v1 Announce Type: new Abstract: Large language models (LLMs) increasingly rely on knowledge editing to support knowledge-intensive reasoning, but this flexibility also introduces criti

model-releasesarxiv-cs-ai
12 May 2026
Safety

Break the Brake, Not the Wheel: Untargeted Jailbreak via Entropy Maximization

DGX agent

arXiv:2605.10764v1 Announce Type: cross Abstract: Recent studies show that gradient-based universal image jailbreaks on vision-language models (VLMs) exhibit little or no cross-model transferability,

safetyarxiv-cs-ai
12 May 2026
Model Releases

CADBench: A Multimodal Benchmark for AI-Assisted CAD Program Generation

DGX agent

arXiv:2605.10873v1 Announce Type: cross Abstract: Recovering editable CAD programs from images or 3D observations is central to AI-assisted design, but progress is difficult to measure because existin

model-releasesarxiv-cs-ai
12 May 2026
Safety

CAMAL: Improving Attention Alignment and Faithfulness with Segmentation Masks

DGX agent

arXiv:2605.08325v1 Announce Type: cross Abstract: Many vision datasets now provide segmentation masks in addition to annotated images to support a wide range of tasks. In this work, we propose Class A

safetyarxiv-cs-ai
12 May 2026
Safety

Can LLMs Estimate Student Struggles? Human-AI Difficulty Alignment with Proficiency Simulation for Item Difficulty Prediction

DGX agent

arXiv:2512.18880v2 Announce Type: replace-cross Abstract: Accurate estimation of item (question or task) difficulty is critical for educational assessment but suffers from the cold start problem. Whil

safetyarxiv-cs-ai
12 May 2026
Safety

Can Revealed Preferences Clarify LLM Alignment and Steering?

DGX agent

arXiv:2605.08556v1 Announce Type: new Abstract: LLMs are increasingly used to make or support high-stakes decisions under uncertainty, where alignment depends not only on factual accuracy but on how m

safetyarxiv-cs-lg
12 May 2026
Model Releases

Causal Stories from Sensor Traces: Auditing Epistemic Overreach in LLM-Generated Personal Sensing Explanations

DGX agent

arXiv:2605.08590v1 Announce Type: cross Abstract: LLMs are increasingly used to explain personal sensing data, translating traces of activity and mood into natural-language accounts of why an anomalou

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

CHAINTRIX: A multi-pipeline LLM-augmented framework for automated smart-contract security auditing

DGX agent

arXiv:2605.09350v1 Announce Type: new Abstract: Smart-contract exploits have caused billions of USD in cumulative losses, yet audits remain expensive and slow. Automated tools have emerged to close th

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

ChatbotManip: A Dataset to Facilitate Evaluation and Oversight of Manipulative Chatbot Behaviour

DGX agent

arXiv:2506.12090v2 Announce Type: replace Abstract: This paper introduces ChatbotManip, a novel dataset for studying manipulation in Chatbots. It contains simulated generated conversations between a c

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Classification-Head Bias in Class-Level Machine Unlearning: Diagnosis, Mitigation, and Evaluation

DGX agent

arXiv:2605.08730v1 Announce Type: new Abstract: Class-level machine unlearning aims to remove the influence of specified classes while preserving model utility on retained classes. Existing methods ar

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

CMKL: Modality-Aware Continual Learning for Evolving Biomedical Knowledge Graphs

DGX agent

arXiv:2605.10510v1 Announce Type: cross Abstract: Biomedical knowledge graphs are increasingly large, dynamic, and multimodal, driven by rapid advances in biotechnology such as high-throughput sequenc

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

CodeClinic: Evaluating Automation of Coding Skills for Clinical Reasoning Agents

DGX agent

arXiv:2605.09675v1 Announce Type: new Abstract: Clinical reasoning agents based on large language models (LLMs) aim to automate tasks such as intensive care unit (ICU) monitoring and patient state tra

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Confidence-Guided Diffusion Augmentation for Enhanced Bangla Compound Character Recognition

DGX agent

arXiv:2605.10916v1 Announce Type: cross Abstract: Recognition of handwritten Bangla compound characters remains a challenging problem due to complex character structures, large intra-class variation,

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Context-Augmented Code Generation: How Product Context Improves AI Coding Agent Decision Compliance by 49%

DGX agent

arXiv:2605.08112v1 Announce Type: cross Abstract: AI coding agents powered by large language models can read codebases and produce functional code, but they routinely violate team-specific product dec

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

CrackMeBench: Binary Reverse Engineering for Agents

DGX agent

arXiv:2605.10597v1 Announce Type: cross Abstract: Benchmarks for coding agents increasingly measure source-level software repair, and cybersecurity benchmarks increasingly measure broad capture-the-fl

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

CTQWformer: A CTQW-based Transformer for Graph Classification

DGX agent

arXiv:2605.09486v1 Announce Type: cross Abstract: Graph Neural Networks (GNN) and Transformer-based architectures have achieved remarkable progress in graph learning, yet they still struggle to captur

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

DECO: Sparse Mixture-of-Experts with Dense-Comparable Performance on End-Side Devices

DGX agent

arXiv:2605.10933v1 Announce Type: cross Abstract: While Mixture-of-Experts (MoE) scales model capacity without proportionally increasing computation, its massive total parameter footprint creates sign

model-releasesarxiv-cs-cl
12 May 2026
Tutorials

Deep Arguing

DGX agent

arXiv:2605.10569v1 Announce Type: new Abstract: Deep learning has become the dominant approach for creating high capacity, scalable models across diverse data modalities. However, because these models

tutorialsarxiv-cs-ai
12 May 2026
Model Releases

Deepfake Detection that Generalizes Across Benchmarks

DGX agent

arXiv:2508.06248v4 Announce Type: replace Abstract: The generalization of deepfake detectors to unseen manipulation techniques remains a challenge for practical deployment. Although many approaches ad

model-releasesarxiv-cs-cv
12 May 2026
Applications

Diagnosing and Mitigating Domain Shift in Permission-Based Android Malware Detection

DGX agent

arXiv:2605.09028v1 Announce Type: new Abstract: Machine learning-based Android malware detectors often fail in real-world deployment due to domain shift, where models trained on one data source perfor

applicationsarxiv-cs-lg
12 May 2026
Model Releases

Distributional Spectral Diagnostics for Localizing Grokking Transitions

DGX agent

arXiv:2605.08237v1 Announce Type: new Abstract: In grokking, a model first fits the training data while test accuracy remains low, and only later begins to generalize. We ask whether this transition c

model-releasesarxiv-cs-lg
12 May 2026
Safety

Do Linear Probes Generalize Better in Persona Coordinates?

DGX agent

arXiv:2605.09391v1 Announce Type: new Abstract: It is becoming increasingly necessary to have monitors check for harmful behaviors during language model interactions, but text-only monitoring has not

safetyarxiv-cs-ai
12 May 2026
Applications

Dolphin-CN-Dialect: Where Chinese Dialects Matter

DGX agent

arXiv:2605.08961v1 Announce Type: new Abstract: We present Dolphin-CN-Dialect, a streaming-capable ASR model with a focus on Chinese and dialect-rich scenarios. Compared to the previous version, Dolph

applicationsarxiv-cs-cl
12 May 2026
Model Releases

Done, But Not Sure: Disentangling World Completion from Self-Termination in Embodied Agents

DGX agent

arXiv:2605.08747v1 Announce Type: new Abstract: Standard embodied evaluations do not independently score whether an agent correctly commits to task completion at episode closure, a capacity we call te

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Don't Click That: Teaching Web Agents to Resist Deceptive Interfaces

DGX agent

arXiv:2605.09497v1 Announce Type: new Abstract: Vision-language model (VLM) based web agents demonstrate impressive autonomous GUI interaction but remain vulnerable to deceptive interface elements. Ex

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Edge-specific signal propagation on mature chromophore-region 3D mechanism graphs for fluorescent protein quantum-yield prediction

DGX agent

arXiv:2605.06644v2 Announce Type: replace Abstract: Fluorescent protein quantum yield (QY) is governed by the mature chromophore and its three-dimensional microenvironment rather than sequence identit

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Efficient Ensemble Selection from Binary and Pairwise Feedback

DGX agent

arXiv:2605.09588v1 Announce Type: cross Abstract: Organizations increasingly deploy multiple AI systems across task domains, but selecting a small, high-performing ensemble can require costly model ca

model-releasesarxiv-cs-ai
12 May 2026
← Previous
1…550551552553554…1394
Next →