AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,975 results
Research

Attention-Based Sampler for Diffusion Language Models

DGX agent

arXiv:2604.08564v1 Announce Type: new Abstract: Auto-regressive models (ARMs) have established a dominant paradigm in language modeling. However, their strictly sequential decoding paradigm imposes fu

researcharxiv-cs-cl
13 Apr 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Benchmarking CNN- and Transformer-Based Models for Surgical Instrument Segmentation in Robotic-Assisted Surgery

DGX agent

arXiv:2604.09151v1 Announce Type: new Abstract: Accurate segmentation of surgical instruments in robotic-assisted surgery is critical for enabling context-aware computer-assisted interventions, such a

model-releasesarxiv-cs-cv
13 Apr 2026
Research

Do Vision Language Models Need to Process Image Tokens?

DGX agent

arXiv:2604.09425v1 Announce Type: new Abstract: Vision Language Models (VLMs) have achieved remarkable success by integrating visual encoders with large language models (LLMs). While VLMs process dens

researcharxiv-cs-cv
13 Apr 2026
Model Releases

Improving Automatic Summarization of Radiology Reports through Mid-Training of Large Language Models

DGX agent

arXiv:2603.19275v2 Announce Type: replace-cross Abstract: Automatic summarization of radiology reports is an essential application to reduce the burden on physicians. Previous studies have widely used

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

AHCQ-SAM: Toward Accurate and Hardware-Compatible Post-Training Segment Anything Model Quantization

DGX agent

arXiv:2503.03088v4 Announce Type: replace-cross Abstract: The Segment Anything Model (SAM) has revolutionized image and video segmentation with its powerful zero-shot capabilities. However, its massiv

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

Emotion Concepts and their Function in a Large Language Model

DGX agent

arXiv:2604.07729v1 Announce Type: cross Abstract: Large language models (LLMs) sometimes appear to exhibit emotional reactions. We investigate why this is the case in Claude Sonnet 4.5 and explore imp

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Entropy After </Think> for reasoning model early exiting

DGX agent

arXiv:2509.26522v3 Announce Type: replace Abstract: Reasoning LLMs show improved performance with longer chains of thought. However, recent work has highlighted their tendency to overthink, continuing

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

FlowGuard: Towards Lightweight In-Generation Safety Detection for Diffusion Models via Linear Latent Decoding

DGX agent

arXiv:2604.07879v1 Announce Type: new Abstract: Diffusion-based image generation models have advanced rapidly but pose a safety risk due to their potential to generate Not-Safe-For-Work (NSFW) content

model-releasesarxiv-cs-cv
10 Apr 2026
Research

From Classical Machine Learning to Tabular Foundation Models: An Empirical Investigation of Robustness and Scalability Under Class Imbalance in Emergency and Critical Care

DGX agent

arXiv:2512.21602v2 Announce Type: replace-cross Abstract: Millions of patients pass through emergency departments and intensive care units each year, where clinicians must make high-stakes decisions u

researcharxiv-cs-cv
10 Apr 2026
Model Releases

Illocutionary Explanation Planning for Source-Faithful Explanations in Retrieval-Augmented Language Models

DGX agent

arXiv:2604.06211v1 Announce Type: cross Abstract: Natural language explanations produced by large language models (LLMs) are often persuasive, but not necessarily scrutable: users cannot easily verify

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

MM-MoralBench: A MultiModal Moral Evaluation Benchmark for Large Vision-Language Models

DGX agent

arXiv:2412.20718v2 Announce Type: replace Abstract: The rapid integration of Large Vision-Language Models (LVLMs) into critical domains necessitates comprehensive moral evaluation to ensure their alig

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

TalkLoRA: Communication-Aware Mixture of Low-Rank Adaptation for Large Language Models

DGX agent

arXiv:2604.06291v1 Announce Type: cross Abstract: Low-Rank Adaptation (LoRA) enables parameter-efficient fine-tuning of Large Language Models (LLMs), and recent Mixture-of-Experts (MoE) extensions fur

model-releasesarxiv-cs-ai
10 Apr 2026
Agents

TOOLCAD: Exploring Tool-Using Large Language Models in Text-to-CAD Generation with Reinforcement Learning

DGX agent

arXiv:2604.07960v1 Announce Type: cross Abstract: Computer-Aided Design (CAD) is an expert-level task that relies on long-horizon reasoning and coherent modeling actions. Large Language Models (LLMs)

agentsarxiv-cs-cl
10 Apr 2026
Safety

BEST-KAG: Enhancing Question Answering of Building Engineering Standards with Multimodal Knowledge Graph Modeling and Large Language Model

DGX agent

arXiv:2608.11244v1 Announce Type: new Abstract: Construction standards are critical for building safety and sustainability. Existing standard application workflows rely on keyword-based document retri

safetyarxiv-cs-ai
13 Aug 2026
Model Releases

Do You See What You Draw? A Semantic Closed-Loop Framework for Holistic Evaluation of Unified Multimodal Models

DGX agent

arXiv:2608.11907v1 Announce Type: cross Abstract: As Large Vision-Language Models increasingly aim to integrate visual generation and understanding within a single parameter space, evaluating such str

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Locating and Controlling Implicit Personalization in Large Language Models

DGX agent

arXiv:2608.11735v1 Announce Type: cross Abstract: Large language models (LLMs) often shift their outputs in response to implicit demographic cues even when users never state a demographic identity. Pr

model-releasesarxiv-cs-ai
13 Aug 2026
Safety

Post-Training with Policy Gradients: Optimality and the Base Model Barrier

DGX agent

arXiv:2603.06957v2 Announce Type: replace-cross Abstract: We study post-training linear autoregressive models with outcome and process rewards. Given a context oldsymbol{x}, the model must predict the

safetyarxiv-cs-ai
13 Aug 2026
Model Releases

Reliable Inference in Edge-Cloud Model Cascades via Conformal Alignment

DGX agent

arXiv:2510.17543v3 Announce Type: replace Abstract: Edge intelligence enables low-latency inference via compact on-device models, but assuring reliability remains challenging. We study edge-cloud casc

model-releasesarxiv-cs-lg
13 Aug 2026
Model Releases

Repurposing RGB-based Foundation Model for Depth Estimation on Thermal Images Using Hierarchical Supervision

DGX agent

arXiv:2608.11564v1 Announce Type: new Abstract: Depth estimation from thermal images is highly valuable for robotic applications in adverse conditions, such as nighttime and rainy weather. Recent stud

model-releasesarxiv-cs-cv
13 Aug 2026
Model Releases

RoboHarness: A Memory-Augmented Policy Harness for Vision-Language-Action Model Robustness via In-Context Adaptation

DGX agent

arXiv:2603.24060v3 Announce Type: replace Abstract: Despite the promise of Vision-Language-Action (VLA) models as generalist robotic controllers, their robustness against perceptual noise and environm

model-releasesarxiv-cs-ro
13 Aug 2026
Applications

Robust Ambiguity Detection (RAD) From Model- and Feature-Space Consistency

DGX agent

arXiv:2608.11541v1 Announce Type: new Abstract: Machine learning models should be robust, in the sense of remaining predictively consistent under permissible variations. A model's predictions should i

applicationsarxiv-cs-lg
13 Aug 2026
Research

Task- and dataset-specific information in protein language models

DGX agent

arXiv:2608.12090v1 Announce Type: new Abstract: Protein language models (PLMs) have transferred the latest advances from natural language processing to computational biology. These models, trained on

researcharxiv-cs-lg
13 Aug 2026
Safety

Do AI weather models miss extremes?

DGX agent

arXiv:2608.09972v1 Announce Type: cross Abstract: First-generation AI weather models are often reported to underperform at extremes, mostly in reanalysis-based evaluations of deterministic regression

safetyarxiv-cs-ai
12 Aug 2026
Research

DoseBridge: Denoising Diffusion Bridge Model for Dose Prediction in Lung Intensity-Modulated Proton Therapy

DGX agent

arXiv:2608.10173v1 Announce Type: new Abstract: Most radiotherapy dose-prediction models use only CT images and anatomical structures, although intensity-modulated proton therapy (IMPT) dose also depe

researcharxiv-cs-cv
12 Aug 2026
Safety

Dreamer-SAC: Off-Policy Learning in Latent World Models for Sample-Efficient Autonomous Driving

DGX agent

arXiv:2608.10386v1 Announce Type: new Abstract: Sample-efficient reinforcement learning for autonomous driving is often limited by the trade-off between data efficiency and model bias. While world mod

safetyarxiv-cs-lg
12 Aug 2026
Research

MoE Proxy Models for Low-Cost Failure Reproduction and Diagnosis in LLM RL Post-Training

DGX agent

arXiv:2608.10823v1 Announce Type: new Abstract: Reinforcement learning (RL) post-training of large language models (LLMs) is computationally intensive and involves complex system pipelines with substa

researcharxiv-cs-lg
12 Aug 2026
Model Releases

ChronoState: Hidden Elapsed-Time Conditioning for Temporal-State Action Selection in Frozen-Backbone Language Models

DGX agent

arXiv:2608.09124v1 Announce Type: new Abstract: Temporal decisions in language-model systems often depend on both symbolic task state and elapsed wall-clock time, such as cache expiration, job complet

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

ComplexityWorld: Benchmarking Vision-Language Models on Verifiable Visual Decision Making

DGX agent

arXiv:2608.07584v1 Announce Type: new Abstract: Vision-language models (VLMs) have made rapid progress in visual perception and increasingly support real-world tasks that depend on images. Many such t

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

EmoS: A Theory-Grounded Framework for Evaluating and Aligning Emotional Intelligence in Spoken Language Models

DGX agent

arXiv:2608.09189v1 Announce Type: new Abstract: Despite significant advances in instruction-following and auditory comprehension, the evaluation of Emotional Intelligence (EI) in Spoken Language Model

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

FoMoH: A clinically meaningful foundation model evaluation for structured electronic health records

DGX agent

arXiv:2505.16941v4 Announce Type: replace-cross Abstract: Foundation models (FMs) promise to address core limitations of traditional supervised machine learning: (i) reliance on large amounts of label

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

In-Loop Model Adaptation with Coupled Latent-Noise Guidance for High-Fidelity Subject-Driven Text-to-Image Generation

DGX agent

arXiv:2608.09244v1 Announce Type: new Abstract: Text-to-image diffusion models have achieved remarkable success in generating high-quality images from a given text prompt. Subject-driven generation ai

safetyarxiv-cs-cv
11 Aug 2026
Model Releases

MathShikkha: A Controlled Study of Answer-Only and Chain-of-Thought Supervision for Bangla Mathematical Reasoning in Small Language Models

DGX agent

arXiv:2608.08503v1 Announce Type: new Abstract: Mathematical reasoning remains challenging in low-resource languages such as Bangla. We study whether teacher-generated Bangla Chain-of-Thought (CoT) su

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

MedPixel: A Unified Pixel-Language Model for Medical Reasoning and Segmentation

DGX agent

arXiv:2608.09818v1 Announce Type: cross Abstract: Reliable medical image understanding requires models to connect clinical language and visual reasoning with pixel-level grounding. Yet medical vision-

model-releasesarxiv-cs-ai
11 Aug 2026
Research

Quantization Degradation in Large Language Models: A Signal-Noise Perspective

DGX agent

arXiv:2608.08188v1 Announce Type: new Abstract: Post-training quantization reduces the deployment cost of large language models, yet how severely a quantized model degrades is not determined by bit-wi

researcharxiv-cs-ai
11 Aug 2026
Model Releases

SDDBMs: Soft Denoising Diffusion Bridge Models

DGX agent

arXiv:2608.08594v1 Announce Type: new Abstract: Diffusion bridge models leverage Doob's (h)-transform to construct stochastic transports between arbitrary endpoint distributions, and have shown strong

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Sekai2: From World Exploration to Interactive World Modeling

DGX agent

arXiv:2608.09449v1 Announce Type: new Abstract: Video world models must capture how scenes evolve over time and across viewpoints. Training them for long-horizon generation and camera control therefor

model-releasesarxiv-cs-cv
11 Aug 2026
Hardware

verdi: retrieval is not transfer for continual world model optimization

DGX agent

arXiv:2608.09537v1 Announce Type: new Abstract: Foundation world models have made remarkable progress in planning, simulation, and embodied intelligence. However, optimizing a pretrained world model t

hardwarearxiv-cs-ai
11 Aug 2026
Safety

World Tokens: Enhancing Embodied Policies with Training-Time World Modeling

DGX agent

arXiv:2608.09730v1 Announce Type: new Abstract: Vision-language-action (VLA) models are a widely adopted paradigm for embodied policies. They excel at efficient closed-loop control but do not explicit

safetyarxiv-cs-cv
11 Aug 2026
Model Releases

WuYuEval: A Multi-Level Benchmark for Large Language Models in Solid Waste Management

DGX agent

arXiv:2608.07529v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as technical assistants, but their competence in solid waste management (SWM) remains difficult to

model-releasesarxiv-cs-ai
11 Aug 2026
Research

ZetaGPT: A Reference Implementation of Positional--Encoding--Free State--Space--Attention Language Models

DGX agent

arXiv:2608.09432v1 Announce Type: cross Abstract: Transformer-based language models rely on self-attention, whose computation is permutation-equivariant and therefore lacks an intrinsic mechanism for

researcharxiv-cs-ai
11 Aug 2026
Research

AfriNLLB: Efficient Translation Models for African Languages

DGX agent

arXiv:2602.09373v2 Announce Type: replace Abstract: In this work, we present AfriNLLB, a series of lightweight models for efficient translation from and into African languages. AfriNLLB supports 15 la

researcharxiv-cs-cl
10 Aug 2026
Model Releases

Ask-E: An Environment for Calibrated Question Generation

DGX agent

arXiv:2608.06933v1 Announce Type: cross Abstract: Today, we improve models by training and evaluating them on problems at the frontier of their abilities. Creating such problems is itself a demanding

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Beyond the Black Box: Interpretable Models of Human Randomisation Failures

DGX agent

arXiv:2608.07220v1 Announce Type: new Abstract: Mixed strategy equilibrium predicts i.i.d play: past actions should not help predict future decisions. Human players, however, systematically depart fro

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Control-Anchored Residual Flow Matching Conditioned on Gene Geometry for Virtual Cell Perturbation Modeling

DGX agent

arXiv:2608.06824v1 Announce Type: cross Abstract: A central task in virtual cell modeling is predicting single-cell transcriptional responses to unseen genetic perturbations and drug combinations, and

model-releasesarxiv-cs-ai
10 Aug 2026
Research

Evaluating Useful Surrogate Models for Configuration Tuning Beyond Accuracy: A Fitness Landscape Analysis Perspective

DGX agent

arXiv:2509.21945v2 Announce Type: replace-cross Abstract: To efficiently tune configuration for better software system performance (e.g., latency) at the deployment and maintenance stage, many tuners

researcharxiv-cs-ai
10 Aug 2026
Safety

Explore or Converge? Stage-Guided Per-Step Optimization for Diffusion Models

DGX agent

arXiv:2608.06768v1 Announce Type: new Abstract: Diffusion models have strong generative capabilities. However, their maximum likelihood training objective only focuses on reconstructing the data distr

safetyarxiv-cs-cv
10 Aug 2026
Model Releases

Index SLM Technical Report

DGX agent

arXiv:2607.09885v2 Announce Type: replace Abstract: We present Index-1.9B, a series of open small language models developed at Bilibili. The series comprises four models: Index-1.9B-Base, a foundation

model-releasesarxiv-cs-cl
10 Aug 2026
Model Releases

Retrofitting Linear Attention into Diffusion Language Models

DGX agent

arXiv:2608.06628v1 Announce Type: new Abstract: Diffusion language models (dLLMs) offer a promising alternative to autoregressive models by accelerating inference through parallel decoding. Recent dLL

model-releasesarxiv-cs-lg
10 Aug 2026
← Previous
1…4142434445…1021
Next →