AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
HumanDGX agent

88,246Total entries
1Added by human
88,245Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,499 results
13 May 2026

Mitigating Context-Memory Conflicts in LLMs through Dynamic Cognitive Reconciliation Decoding

Model ReleasesDGX agent

arXiv:2605.12185v1 Announce Type: new Abstract: Large language models accumulate extensive parametric knowledge through pre-training. However, knowledge conflicts occur when outdated or incorrect para

MULTI: Disentangling Camera Lens, Sensor, View, and Domain for Novel Image Generation

Model ReleasesDGX agent

arXiv:2605.12134v1 Announce Type: new Abstract: Recent text-to-image models produce high-quality images, yet text ambiguity hinders precise control when specific styles or objects are required. There

Not How Many, But Which: Parameter Placement in Low-Rank Adaptation

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.12207v1 Announce Type: cross Abstract: We study the extit{parameter placement problem}: given a fixed budget of k trainable entries within the B matrix of a LoRA adapter (A frozen), does th

ORBIT: Preserving Foundational Language Capabilities in GenRetrieval via Origin-Regulated Merging

ResearchDGX agent

arXiv:2605.12419v1 Announce Type: new Abstract: Despite the rapid advancements in large language model (LLM) development, fine-tuning them for specific tasks often results in the catastrophic forgetti

Paper: http://arxiv.org/abs/2605.06546 HF: http://huggingface.co/papers/2605.06546 Blog: http://nousresearch.com/token-superposition

ResearchDGX agent

This paper investigates token superposition, a phenomenon where language models can encode multiple token representations simultaneously in a single position, enabling more efficient use of model capa

PD-4DGS:Progressive Decomposition of 4D Gaussian Splatting for Bandwidth-Adaptive Dynamic Scene Streaming

Model ReleasesDGX agent

arXiv:2605.11427v1 Announce Type: new Abstract: 4D Gaussian Splatting (4DGS) enables high-quality dynamic novel view synthesis, yet current models remain monolithic bitstreams that clients must downlo

Position: Universal Aesthetic Alignment Narrows Artistic Expression

SafetyDGX agent

arXiv:2512.11883v3 Announce Type: replace-cross Abstract: Over-aligning image generation models to a generalized aesthetic preference conflicts with user intent, particularly when 'anti-aesthetic' out

Prompting from the bench: Large-scale pretraining is not sufficient to prepare LLMs for ordinary meaning analysis

ApplicationsDGX agent

arXiv:2510.25356v2 Announce Type: replace Abstract: In the U.S. judicial system, a widespread approach to legal interpretation entails assessing how a legal text would be understood by an `ordinary' s

Provably Data-driven Multiple Hyper-parameter Tuning with Structured Loss Function

Model ReleasesDGX agent

arXiv:2602.02406v2 Announce Type: replace-cross Abstract: Data-driven algorithm design automates hyperparameter tuning, but its statistical foundations remain limited because model performance can dep

Revisiting Shadow Detection from a Vision-Language Perspective

Model ReleasesDGX agent

arXiv:2605.11771v1 Announce Type: new Abstract: Shadow detection is commonly formulated as a vision-driven dense prediction problem, where models rely primarily on pixel-wise visual supervision to dis

SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture

AgentsDGX agent

arXiv:2605.12500v1 Announce Type: new Abstract: Recent large vision-language models (VLMs) remain fundamentally constrained by a persistent dichotomy: understanding and generation are treated as disti

ShapeCodeBench: A Renewable Benchmark for Perception-to-Program Reconstruction of Synthetic Shape Scenes

Model ReleasesDGX agent

arXiv:2605.11680v1 Announce Type: new Abstract: We introduce ShapeCodeBench, a synthetic benchmark for perception-to-program reconstruction: given a rendered raster image, a model must emit an executa

Targeted Tests for LLM Reasoning: An Audit-Constrained Protocol

ResearchDGX agent

arXiv:2605.11599v1 Announce Type: new Abstract: Fixed reasoning benchmarks evaluate canonical prompts, but semantically valid changes in presentation can still change model behavior. Studies of prompt

The new era of SaMD: Why cloud infrastructure is the foundation for digital health in 2026

SafetyDGX agent

In the healthcare and life sciences industries, speed saves lives, but meeting regulatory requirements and other administrative burdens often pumps the brakes for manufacturers of software as a medica

The Price of Proportional Representation in Temporal Voting

Model ReleasesDGX agent

arXiv:2605.11157v1 Announce Type: cross Abstract: We study proportional representation in the temporal voting model, where collective decisions are made repeatedly over time over a fixed horizon. Prio

Urban Risk-Aware Navigation via VQA-Based Event Maps for People with Low Vision

Model ReleasesDGX agent

arXiv:2605.11782v1 Announce Type: new Abstract: Visual impairment affects hundreds of millions of people worldwide, severely limiting their ability to navigate urban environments safely and independen

Will Ollama come out with a non-cloud version of Deepseek-v4 Flash?

Model ReleasesDGX agent

DeepSeek-v4 Flash through Ollama is currently available as a cloud model, where Ollama's CLI sends API calls to Ollama's hosted version rather than running locally . Local support for DeepSeek V4 Flas

12 May 2026

A Cognitively Grounded Bayesian Framework for Misinformation Susceptibility

Model ReleasesDGX agent

arXiv:2605.09483v1 Announce Type: cross Abstract: In this (work in progress) paper, we present Bounded Pragmatic Listener (or BPL), a cognitively grounded Bayesian framework for modelling susceptibili

A Qualitative Test-Risk Mechanism for Scaling Behavior in Normalized Residual Networks

ResearchDGX agent

arXiv:2605.08297v1 Announce Type: cross Abstract: The scaling behavior, in which test performance often improves as model size and data increase, is a central empirical phenomenon in modern deep learn

A Scalable Entity-Based Framework for Auditing Bias in LLMs

SafetyDGX agent

arXiv:2601.12374v2 Announce Type: replace-cross Abstract: Existing approaches to bias evaluation in large language models (LLMs) trade ecological validity for statistical control, relying either on ar

Action-Guided Attention for Video Action Anticipation

Model ReleasesDGX agent

arXiv:2603.01743v2 Announce Type: replace Abstract: Anticipating future actions in videos is challenging, as the observed frames provide only evidence of past activities, requiring the inference of la

AdaPaD: Adaptive Parallel Deflation for PEFT with Self-Correcting Rank Discovery

Model ReleasesDGX agent

arXiv:2605.10741v1 Announce Type: new Abstract: Fine-tuning large language models with LoRA requires choosing a rank r before training starts. Existing approaches either extract rank-1 components sequ

Agent-ValueBench: A Comprehensive Benchmark for Evaluating Agent Values

Model ReleasesDGX agent

arXiv:2605.10365v1 Announce Type: new Abstract: Autonomous agents have rapidly matured as task executors and seen widespread deployment via harnesses such as OpenClaw. Safety concerns have rightly dra

An Annotation Scheme and Classifier for Personal Facts in Dialogue

Model ReleasesDGX agent

arXiv:2605.10339v1 Announce Type: new Abstract: The advancement of Large Language Models (LLMs) has enabled their application in personalized dialogue systems. We present an extended annotation scheme

An Empirical Study of Multi-Agent Collaboration for Automated Research

Model ReleasesDGX agent

arXiv:2603.29632v2 Announce Type: replace-cross Abstract: As AI agents evolve, the community is rapidly shifting from single Large Language Models (LLMs) to Multi-Agent Systems (MAS) to overcome cogni

Annotations Mitigate Post-Training Mode Collapse

TutorialsDGX agent

arXiv:2605.09995v1 Announce Type: new Abstract: Post-training (via supervised fine-tuning) improves instruction-following, but often induces semantic mode collapse by biasing models toward low-entropy

AnyDepth-DETR/-YOLO: Any-depth object detection with a single network

Model ReleasesDGX agent

arXiv:2605.09407v1 Announce Type: new Abstract: Modern object detectors are static, fixed-depth networks optimized for a single operating point, requiring separate models for different deployment scen

AssayBench: An Assay-Level Virtual Cell Benchmark for LLMs and Agents

Model ReleasesDGX agent

arXiv:2605.10876v1 Announce Type: cross Abstract: Recent advances in machine learning and large-scale biological data collections have revived the prospect of building a virtual cell, a computational

Attention Grounded Enhancement for Visual Document Retrieval

Model ReleasesDGX agent

arXiv:2511.13415v2 Announce Type: replace-cross Abstract: Visual document retrieval requires understanding heterogeneous and multi-modal content to satisfy implicit information needs. Recent advances

AUHead: Realistic Emotional Talking Head Generation via Action Units Control

Model ReleasesDGX agent

arXiv:2602.09534v2 Announce Type: replace Abstract: Realistic talking-head video generation is critical for virtual avatars, film production, and interactive systems. Current methods struggle with nua

Benchmarking Safety Risks of Knowledge-Intensive Reasoning under Malicious Knowledge Editing

Model ReleasesDGX agent

arXiv:2605.10146v1 Announce Type: new Abstract: Large language models (LLMs) increasingly rely on knowledge editing to support knowledge-intensive reasoning, but this flexibility also introduces criti

Break the Brake, Not the Wheel: Untargeted Jailbreak via Entropy Maximization

SafetyDGX agent

arXiv:2605.10764v1 Announce Type: cross Abstract: Recent studies show that gradient-based universal image jailbreaks on vision-language models (VLMs) exhibit little or no cross-model transferability,

CADBench: A Multimodal Benchmark for AI-Assisted CAD Program Generation

Model ReleasesDGX agent

arXiv:2605.10873v1 Announce Type: cross Abstract: Recovering editable CAD programs from images or 3D observations is central to AI-assisted design, but progress is difficult to measure because existin

CAMAL: Improving Attention Alignment and Faithfulness with Segmentation Masks

SafetyDGX agent

arXiv:2605.08325v1 Announce Type: cross Abstract: Many vision datasets now provide segmentation masks in addition to annotated images to support a wide range of tasks. In this work, we propose Class A

Can LLMs Estimate Student Struggles? Human-AI Difficulty Alignment with Proficiency Simulation for Item Difficulty Prediction

SafetyDGX agent

arXiv:2512.18880v2 Announce Type: replace-cross Abstract: Accurate estimation of item (question or task) difficulty is critical for educational assessment but suffers from the cold start problem. Whil

Can Revealed Preferences Clarify LLM Alignment and Steering?

SafetyDGX agent

arXiv:2605.08556v1 Announce Type: new Abstract: LLMs are increasingly used to make or support high-stakes decisions under uncertainty, where alignment depends not only on factual accuracy but on how m

Causal Stories from Sensor Traces: Auditing Epistemic Overreach in LLM-Generated Personal Sensing Explanations

Model ReleasesDGX agent

arXiv:2605.08590v1 Announce Type: cross Abstract: LLMs are increasingly used to explain personal sensing data, translating traces of activity and mood into natural-language accounts of why an anomalou

CHAINTRIX: A multi-pipeline LLM-augmented framework for automated smart-contract security auditing

Model ReleasesDGX agent

arXiv:2605.09350v1 Announce Type: new Abstract: Smart-contract exploits have caused billions of USD in cumulative losses, yet audits remain expensive and slow. Automated tools have emerged to close th

ChatbotManip: A Dataset to Facilitate Evaluation and Oversight of Manipulative Chatbot Behaviour

Model ReleasesDGX agent

arXiv:2506.12090v2 Announce Type: replace Abstract: This paper introduces ChatbotManip, a novel dataset for studying manipulation in Chatbots. It contains simulated generated conversations between a c

Classification-Head Bias in Class-Level Machine Unlearning: Diagnosis, Mitigation, and Evaluation

Model ReleasesDGX agent

arXiv:2605.08730v1 Announce Type: new Abstract: Class-level machine unlearning aims to remove the influence of specified classes while preserving model utility on retained classes. Existing methods ar

CMKL: Modality-Aware Continual Learning for Evolving Biomedical Knowledge Graphs

Model ReleasesDGX agent

arXiv:2605.10510v1 Announce Type: cross Abstract: Biomedical knowledge graphs are increasingly large, dynamic, and multimodal, driven by rapid advances in biotechnology such as high-throughput sequenc

CodeClinic: Evaluating Automation of Coding Skills for Clinical Reasoning Agents

Model ReleasesDGX agent

arXiv:2605.09675v1 Announce Type: new Abstract: Clinical reasoning agents based on large language models (LLMs) aim to automate tasks such as intensive care unit (ICU) monitoring and patient state tra

Confidence-Guided Diffusion Augmentation for Enhanced Bangla Compound Character Recognition

Model ReleasesDGX agent

arXiv:2605.10916v1 Announce Type: cross Abstract: Recognition of handwritten Bangla compound characters remains a challenging problem due to complex character structures, large intra-class variation,

Context-Augmented Code Generation: How Product Context Improves AI Coding Agent Decision Compliance by 49%

Model ReleasesDGX agent

arXiv:2605.08112v1 Announce Type: cross Abstract: AI coding agents powered by large language models can read codebases and produce functional code, but they routinely violate team-specific product dec

CrackMeBench: Binary Reverse Engineering for Agents

Model ReleasesDGX agent

arXiv:2605.10597v1 Announce Type: cross Abstract: Benchmarks for coding agents increasingly measure source-level software repair, and cybersecurity benchmarks increasingly measure broad capture-the-fl

CTQWformer: A CTQW-based Transformer for Graph Classification

Model ReleasesDGX agent

arXiv:2605.09486v1 Announce Type: cross Abstract: Graph Neural Networks (GNN) and Transformer-based architectures have achieved remarkable progress in graph learning, yet they still struggle to captur

DECO: Sparse Mixture-of-Experts with Dense-Comparable Performance on End-Side Devices

Model ReleasesDGX agent

arXiv:2605.10933v1 Announce Type: cross Abstract: While Mixture-of-Experts (MoE) scales model capacity without proportionally increasing computation, its massive total parameter footprint creates sign

Deep Arguing

TutorialsDGX agent

arXiv:2605.10569v1 Announce Type: new Abstract: Deep learning has become the dominant approach for creating high capacity, scalable models across diverse data modalities. However, because these models

Deepfake Detection that Generalizes Across Benchmarks

Model ReleasesDGX agent

arXiv:2508.06248v4 Announce Type: replace Abstract: The generalization of deepfake detectors to unseen manipulation techniques remains a challenge for practical deployment. Although many approaches ad

Diagnosing and Mitigating Domain Shift in Permission-Based Android Malware Detection

ApplicationsDGX agent

arXiv:2605.09028v1 Announce Type: new Abstract: Machine learning-based Android malware detectors often fail in real-world deployment due to domain shift, where models trained on one data source perfor

Distributional Spectral Diagnostics for Localizing Grokking Transitions

Model ReleasesDGX agent

arXiv:2605.08237v1 Announce Type: new Abstract: In grokking, a model first fits the training data while test accuracy remains low, and only later begins to generalize. We ask whether this transition c

Do Linear Probes Generalize Better in Persona Coordinates?

SafetyDGX agent

arXiv:2605.09391v1 Announce Type: new Abstract: It is becoming increasingly necessary to have monitors check for harmful behaviors during language model interactions, but text-only monitoring has not

Dolphin-CN-Dialect: Where Chinese Dialects Matter

ApplicationsDGX agent

arXiv:2605.08961v1 Announce Type: new Abstract: We present Dolphin-CN-Dialect, a streaming-capable ASR model with a focus on Chinese and dialect-rich scenarios. Compared to the previous version, Dolph

Done, But Not Sure: Disentangling World Completion from Self-Termination in Embodied Agents

Model ReleasesDGX agent

arXiv:2605.08747v1 Announce Type: new Abstract: Standard embodied evaluations do not independently score whether an agent correctly commits to task completion at episode closure, a capacity we call te

Don't Click That: Teaching Web Agents to Resist Deceptive Interfaces

Model ReleasesDGX agent

arXiv:2605.09497v1 Announce Type: new Abstract: Vision-language model (VLM) based web agents demonstrate impressive autonomous GUI interaction but remain vulnerable to deceptive interface elements. Ex

Edge-specific signal propagation on mature chromophore-region 3D mechanism graphs for fluorescent protein quantum-yield prediction

Model ReleasesDGX agent

arXiv:2605.06644v2 Announce Type: replace Abstract: Fluorescent protein quantum yield (QY) is governed by the mature chromophore and its three-dimensional microenvironment rather than sequence identit

Efficient Ensemble Selection from Binary and Pairwise Feedback

Model ReleasesDGX agent

arXiv:2605.09588v1 Announce Type: cross Abstract: Organizations increasingly deploy multiple AI systems across task domains, but selecting a small, high-performing ensemble can require costly model ca

Elastic MoE: Unlocking the Inference-Time Scalability of Mixture-of-Experts

ApplicationsDGX agent

arXiv:2509.21892v2 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) models typically fix the number of activated experts k at both training and inference. However, real-world deployment

ELF: Embedded Language Flows

ResearchDGX agent

arXiv:2605.10938v1 Announce Type: cross Abstract: Diffusion and flow-based models have become the de facto approaches for generating continuous data, e.g., in domains such as images and videos. Their

ER-Reason: A Benchmark Dataset for LLM Clinical Reasoning in the Emergency Room

Model ReleasesDGX agent

arXiv:2505.22919v3 Announce Type: replace Abstract: Existing benchmarks for evaluating the clinical reasoning capabilities of large language models (LLMs) often lack a clear definition of 'clinical re

← Previous
1…415416417418419…1059
Next →