AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
16,965 results
Model Releases

Break Me If You Can: Self-Jailbreaking of Aligned LLMs via Lexical Insertion Prompting

DGX agent

arXiv:2601.02670v2 Announce Type: replace Abstract: We introduce self-jailbreaking, a threat model in which an aligned LLM guides its own compromise. Unlike most jailbreak techniques, which oft

model-releasesarxiv-cs-cl
10 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Bridging Theory and Practice in Crafting Robust Spiking Reservoirs

DGX agent

arXiv:2604.06395v1 Announce Type: new Abstract: Spiking reservoir computing provides an energy-efficient approach to temporal processing, but reliably tuning reservoirs to operate at the edge-of-chaos

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

Broken by Default: A Formal Verification Study of Security Vulnerabilities in AI-Generated Code

DGX agent

arXiv:2604.05292v2 Announce Type: replace-cross Abstract: AI coding assistants are now used to generate production code in security-sensitive domains, yet the exploitability of their outputs remains u

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

BTC-LLM: Efficient Sub-1-Bit LLM Quantization via Learnable Transformation and Binary Codebook

DGX agent

arXiv:2506.12040v2 Announce Type: replace-cross Abstract: Binary quantization represents the most extreme form of compression, reducing weights to +/-1 for maximal memory and computational efficiency.

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

CADENCE: Context-Adaptive Depth Estimation for Navigation and Computational Efficiency

DGX agent

arXiv:2604.07286v1 Announce Type: cross Abstract: Autonomous vehicles deployed in remote environments typically rely on embedded processors, compact batteries, and lightweight sensors. These hardware

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Calibration of a neural network ocean closure for improved mean state and variability

DGX agent

arXiv:2604.06398v1 Announce Type: cross Abstract: Global ocean models exhibit biases in the mean state and variability, particularly at coarse resolution, where mesoscale eddies are unresolved. To add

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

CAMO: A Class-Aware Minority-Optimized Ensemble for Robust Language Model Evaluation on Imbalanced Data

DGX agent

arXiv:2604.07583v1 Announce Type: new Abstract: Real-world categorization is severely hampered by class imbalance because traditional ensembles favor majority classes, which lowers minority performanc

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

CAMotion: A High-Quality Benchmark for Camouflaged Moving Object Detection in the Wild

DGX agent

arXiv:2604.08287v1 Announce Type: new Abstract: Discovering camouflaged objects is a challenging task in computer vision due to the high similarity between camouflaged objects and their surroundings.

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Can Vision Language Models Judge Action Quality? An Empirical Evaluation

DGX agent

arXiv:2604.08294v1 Announce Type: cross Abstract: Action Quality Assessment (AQA) has broad applications in physical therapy, sports coaching, and competitive judging. Although Vision Language Models

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

CASE: Cadence-Aware Set Encoding for Large-Scale Next Basket Repurchase Recommendation

DGX agent

arXiv:2604.06718v2 Announce Type: cross Abstract: Repurchase behavior is a primary signal in large-scale retail recommendation, particularly in categories with frequent replenishment: many items in a

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

Chunks as Arms: Multi-Armed Bandit-Guided Sampling for Long-Context LLM Preference Optimization

DGX agent

arXiv:2508.13993v2 Announce Type: replace Abstract: Long-context modeling is critical for a wide range of real-world tasks, including long-context question answering, summarization, and complex reason

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

ClawBench: Can AI Agents Complete Everyday Online Tasks?

DGX agent

arXiv:2604.08523v1 Announce Type: new Abstract: AI agents may be able to automate your inbox, but can they automate other routine aspects of your life? Everyday online tasks offer a realistic yet unso

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

ClawsBench: Evaluating Capability and Safety of LLM Productivity Agents in Simulated Workspaces

DGX agent

arXiv:2604.05172v2 Announce Type: replace Abstract: Large language model (LLM) agents are increasingly deployed to automate productivity tasks (e.g., email, scheduling, document management), but evalu

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Cognitive Mismatch in Multimodal Large Language Models for Discrete Symbol Understanding

DGX agent

arXiv:2603.18472v2 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) perform strongly on natural images, yet their ability to understand discrete visual symbols remains u

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Commander-GPT: Dividing and Routing for Multimodal Sarcasm Detection

DGX agent

arXiv:2506.19420v2 Announce Type: replace Abstract: Multimodal sarcasm understanding is a high-order cognitive task. Although large language models (LLMs) have shown impressive performance on many dow

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Consistency-Guided Decoding with Proof-Driven Disambiguation for Three-Way Logical Question Answering

DGX agent

arXiv:2604.06196v1 Announce Type: cross Abstract: Three-way logical question answering (QA) assigns True/False/Unknown to a hypothesis H given a premise set S. While modern large language models

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

ConsistRM: Improving Generative Reward Models via Consistency-Aware Self-Training

DGX agent

arXiv:2604.07484v1 Announce Type: cross Abstract: Generative reward models (GRMs) have emerged as a promising approach for aligning Large Language Models (LLMs) with human preferences by offering grea

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Contextual Earnings-22: A Speech Recognition Benchmark with Custom Vocabulary in the Wild

DGX agent

arXiv:2604.07354v1 Announce Type: new Abstract: The accuracy frontier of speech-to-text systems has plateaued on academic benchmarks.1 In contrast, industrial benchmarks and adoption in high-stakes do

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Continual Visual Anomaly Detection on the Edge: Benchmark and Efficient Solutions

DGX agent

arXiv:2604.06435v1 Announce Type: cross Abstract: Visual Anomaly Detection (VAD) is a critical task for many applications including industrial inspection and healthcare. While VAD has been extensively

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

ConvoLearn: A Dataset for Fine-Tuning Dialogic AI Tutors

DGX agent

arXiv:2601.08950v3 Announce Type: replace Abstract: Despite their growing adoption in education, LLMs remain misaligned with the core principle of effective tutoring: the dialogic construction of know

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

CrashSight: A Phase-Aware, Infrastructure-Centric Video Benchmark for Traffic Crash Scene Understanding and Reasoning

DGX agent

arXiv:2604.08457v1 Announce Type: new Abstract: Cooperative autonomous driving requires traffic scene understanding from both vehicle and infrastructure perspectives. While vision-language models (VLM

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Cross-Lingual Transfer and Parameter-Efficient Adaptation in the Turkic Language Family: A Theoretical Framework for Low-Resource Language Models

DGX agent

arXiv:2604.06202v1 Announce Type: cross Abstract: Large language models (LLMs) have transformed natural language processing, yet their capabilities remain uneven across languages. Most multilingual mo

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

CryoSplat: Gaussian Splatting for Cryo-EM Homogeneous Reconstruction

DGX agent

arXiv:2508.04929v4 Announce Type: replace-cross Abstract: As a critical modality for structural biology, cryogenic electron microscopy (cryo-EM) facilitates the determination of macromolecular structu

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

CycleChart: A Unified Consistency-Based Learning Framework for Bidirectional Chart Understanding and Generation

DGX agent

arXiv:2512.19173v2 Announce Type: replace Abstract: Current chart-related tasks, such as chart generation (NL2Chart), chart schema parsing, chart data parsing, and chart question answering (ChartQA),

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

DDP-SA: Scalable Privacy-Preserving Federated Learning via Distributed Differential Privacy and Secure Aggregation

DGX agent

arXiv:2604.07125v1 Announce Type: cross Abstract: This article presents DDP-SA, a scalable privacy-preserving federated learning framework that jointly leverages client-side local differential privacy

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

Detecting HIV-Related Stigma in Clinical Narratives Using Large Language Models

DGX agent

arXiv:2604.07717v1 Announce Type: new Abstract: Human immunodeficiency virus (HIV)-related stigma is a critical psychosocial determinant of health for people living with HIV (PLWH), influencing mental

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Diagnosing and Mitigating Sycophancy and Skepticism in LLM Causal Judgment

DGX agent

arXiv:2601.08258v3 Announce Type: replace Abstract: Large language models increasingly fail in a way that scalar accuracy cannot diagnose: they produce a sound reasoning trace and then abandon it unde

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Digital Skin, Digital Bias: Uncovering Tone-Based Biases in LLMs and Emoji Embeddings

DGX agent

arXiv:2604.06863v1 Announce Type: cross Abstract: Skin-toned emojis are crucial for fostering personal identity and social inclusion in online communication. As AI models, particularly Large Language

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Direct Segmentation without Logits Optimization for Training-Free Open-Vocabulary Semantic Segmentation

DGX agent

arXiv:2604.07723v1 Announce Type: new Abstract: Open-vocabulary semantic segmentation (OVSS) aims to segment arbitrary category regions in images using open-vocabulary prompts, necessitating that exis

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

DISSECT: Diagnosing Where Vision Ends and Language Priors Begin in Scientific VLMs

DGX agent

arXiv:2604.06250v1 Announce Type: cross Abstract: When asked to describe a molecular diagram, a Vision-Language Model correctly identifies ``a benzene ring with an -OH group.'' When asked to reason ab

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Distributed Interpretability and Control for Large Language Models

DGX agent

arXiv:2604.06483v1 Announce Type: cross Abstract: Large language models that require multiple GPU cards to host are usually the most capable models. It is necessary to understand and steer these model

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Distributed Multi-Layer Editing for Rule-Level Knowledge in Large Language Models

DGX agent

arXiv:2604.08284v1 Announce Type: new Abstract: Large language models store not only isolated facts but also rules that support reasoning across symbolic expressions, natural language explanations, an

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Do MLLMs Really Understand Space? A Mathematical Reasoning Evaluation

DGX agent

arXiv:2602.11635v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have achieved strong performance on perception-oriented tasks, yet their ability to perform mathematical sp

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Domain-Contextualized Inference: A Computable Graph Architecture for Explicit-Domain Reasoning

DGX agent

arXiv:2604.04344v2 Announce Type: replace Abstract: We establish a computation-substrate-agnostic inference architecture in which domain is an explicit first-class computational parameter. This produc

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Don't Overthink It: Inter-Rollout Action Agreement as a Free Adaptive-Compute Signal for LLM Agents

DGX agent

arXiv:2604.08369v1 Announce Type: cross Abstract: Inference-time compute scaling has emerged as a powerful technique for improving the reliability of large language model (LLM) agents, but existing me

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

DosimeTron: Automating Personalized Monte Carlo Radiation Dosimetry in PET/CT with Agentic AI

DGX agent

arXiv:2604.06280v1 Announce Type: cross Abstract: Purpose: To develop and evaluate DosimeTron, an agentic AI system for automated patient-specific MC internal radiation dosimetry in PET/CT examination

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Draw-In-Mind: Rebalancing Designer-Painter Roles in Unified Multimodal Models Benefits Image Editing

DGX agent

arXiv:2509.01986v4 Announce Type: replace-cross Abstract: In recent years, integrating multimodal understanding and generation into a single unified model has emerged as a promising paradigm. While th

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

DROP: Distributional and Regular Optimism and Pessimism for Reinforcement Learning

DGX agent

arXiv:2410.17473v2 Announce Type: replace Abstract: In reinforcement learning (RL), temporal difference (TD) error is known to be related to the firing rate of dopamine neurons. It has been observed t

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

DSCA: Dynamic Subspace Concept Alignment for Lifelong VLM Editing

DGX agent

arXiv:2604.07965v1 Announce Type: new Abstract: Model editing aims to update knowledge to add new concepts and change relevant information without retraining. Lifelong editing is a challenging task, p

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Dual-level Modality Debiasing Learning for Unsupervised Visible-Infrared Person Re-Identification

DGX agent

arXiv:2512.03745v2 Announce Type: replace Abstract: Two-stage learning pipeline has achieved promising results in unsupervised visible-infrared person re-identification (USL-VI-ReID). It first perform

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Dual-Pool Token-Budget Routing for Cost-Efficient and Reliable LLM Serving

DGX agent

arXiv:2604.08075v1 Announce Type: new Abstract: Production vLLM fleets typically provision each instance for the worst-case context length, leading to substantial KV-cache over-allocation and under-ut

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Dynamic Context Evolution for Scalable Synthetic Data Generation

DGX agent

arXiv:2604.07147v1 Announce Type: cross Abstract: Large language models produce repetitive output when prompted independently across many batches, a phenomenon we term cross-batch mode collapse: the p

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

E2Edev: Benchmarking Large Language Models in End-to-End Software Development Task

DGX agent

arXiv:2510.14509v3 Announce Type: replace-cross Abstract: The rapid advancement in large language models (LLMs) has demonstrated significant potential in End-to-End Software Development (E2ESD). Howev

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

EditCaption: Human-Aligned Instruction Synthesis for Image Editing via Supervised Fine-Tuning and Direct Preference Optimization

DGX agent

arXiv:2604.08213v1 Announce Type: new Abstract: High-quality training triplets (source-target image pairs with precise editing instructions) are a critical bottleneck for scaling instruction-guided im

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Efficient and Effective Internal Memory Retrieval for LLM-Based Healthcare Prediction

DGX agent

arXiv:2604.07659v1 Announce Type: new Abstract: Large language models (LLMs) hold significant promise for healthcare, yet their reliability in high-stakes clinical settings is often compromised by hal

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Emotion Concepts and their Function in a Large Language Model

DGX agent

arXiv:2604.07729v1 Announce Type: cross Abstract: Large language models (LLMs) sometimes appear to exhibit emotional reactions. We investigate why this is the case in Claude Sonnet 4.5 and explore imp

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Enabling Intrinsic Reasoning over Dense Geospatial Embeddings with DFR-Gemma

DGX agent

arXiv:2604.07490v1 Announce Type: new Abstract: Representation learning for geospatial and spatio-temporal data plays a critical role in enabling general-purpose geospatial intelligence. Recent geospa

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Entropy After </Think> for reasoning model early exiting

DGX agent

arXiv:2509.26522v3 Announce Type: replace Abstract: Reasoning LLMs show improved performance with longer chains of thought. However, recent work has highlighted their tendency to overthink, continuing

model-releasesarxiv-cs-lg
10 Apr 2026
← Previous
1…347348349350351…354
Next →