AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,543 results
Safety

Characterizing Model-Native Skills

DGX agent

arXiv:2604.17614v1 Announce Type: cross Abstract: Skills are a natural unit for describing what a language model can do and how its behavior can be changed. However, existing characterizations rely on

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Cross-Family Speculative Decoding for Polish Language Models on Apple~Silicon: An Empirical Evaluation of Bielik~11B with UAG-Extended MLX-LM

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2604.16368v1 Announce Type: new Abstract: Speculative decoding accelerates LLM inference by using a small draft model to propose k candidate tokens for a target model to verify. While effective

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Injecting Structured Biomedical Knowledge into Language Models: Continual Pretraining vs. GraphRAG

DGX agent

arXiv:2604.16422v1 Announce Type: new Abstract: The injection of domain-specific knowledge is crucial for adapting language models (LMs) to specialized fields such as biomedicine. While most current a

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

JudgeMeNot: Personalizing Large Language Models to Emulate Judicial Reasoning in Hebrew

DGX agent

arXiv:2604.18041v1 Announce Type: new Abstract: Despite significant advances in large language models, personalizing them for individual decision-makers remains an open problem. Here, we introduce a s

model-releasesarxiv-cs-cl
21 Apr 2026
Research

LogicDiff: Logic-Guided Denoising Improves Zero-Shot Reasoning in Masked Diffusion Language Models

DGX agent

arXiv:2603.26771v2 Announce Type: replace Abstract: Masked diffusion language models (MDLMs) generate text by iteratively unmasking tokens from a fully masked sequence. Their standard confidence-based

researcharxiv-cs-cl
21 Apr 2026
Research

Reasoning Models Know What's Important, and Encode It in Their Activations

DGX agent

arXiv:2604.18307v1 Announce Type: new Abstract: Language models often solve complex tasks by generating long reasoning chains, consisting of many steps with varying importance. While some steps are cr

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Scaling Recurrence-aware Foundation Models for Clinical Records via Next-Visit Prediction

DGX agent

arXiv:2603.24562v2 Announce Type: replace Abstract: While large-scale pretraining has revolutionized language modeling, its potential remains underexplored in healthcare with structured electronic hea

model-releasesarxiv-cs-lg
21 Apr 2026
Safety

Sharpening Lightweight Models for Generalized Polyp Segmentation: A Boundary Guided Distillation from Foundation Models

DGX agent

arXiv:2604.17865v1 Announce Type: new Abstract: Automated polyp segmentation is critical for early colorectal cancer detection and its prevention, yet remains challenging due to weak boundaries, large

safetyarxiv-cs-cv
21 Apr 2026
Model Releases

SIF: Semantically In-Distribution Fingerprints for Large Vision-Language Models

DGX agent

arXiv:2604.17041v1 Announce Type: new Abstract: The public accessibility of large vision-language models (LVLMs) raises serious concerns about unauthorized model reuse and intellectual property infrin

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Systematic Capability Benchmarking of Frontier Large Language Models for Offensive Cyber Tasks

DGX agent

arXiv:2604.17159v1 Announce Type: cross Abstract: We present, to our knowledge, the most comprehensive cross-model evaluation of LLM agents on offensive cybersecurity tasks, benchmarking 10 frontier m

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

The Illusion of Insight in Reasoning Models

DGX agent

arXiv:2601.00514v2 Announce Type: replace-cross Abstract: Do reasoning models have 'Aha!' moments? Prior work suggests that models like DeepSeek-R1-Zero undergo sudden mid-trace realizations that lead

model-releasesarxiv-cs-cl
21 Apr 2026
Tutorials

Towards a Foundation-Model Paradigm for Aerodynamic Prediction in Three-dimensional Design

DGX agent

arXiv:2604.18062v1 Announce Type: new Abstract: Accurate machine-learning models for aerodynamic prediction are essential for accelerating shape optimization, yet remain challenging to develop for com

tutorialsarxiv-cs-lg
21 Apr 2026
Model Releases

VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction

DGX agent

arXiv:2505.20279v4 Announce Type: replace-cross Abstract: The rapid advancement of Large Multimodal Models (LMMs) for 2D images and videos has motivated extending these models to understand 3D scenes,

model-releasesarxiv-cs-cl
21 Apr 2026
Applications

Evaluating the Progression of Large Language Model Capabilities for Small-Molecule Drug Design

DGX agent

arXiv:2604.16279v1 Announce Type: new Abstract: Large Language Models (LLMs) have the potential to accelerate small molecule drug design due to their ability to reason about information from diverse s

applicationsarxiv-cs-lg
20 Apr 2026
Model Releases

Fragile Thoughts: How Large Language Models Handle Chain-of-Thought Perturbations

DGX agent

arXiv:2603.03332v3 Announce Type: replace-cross Abstract: Chain-of-Thought (CoT) prompting has emerged as a foundational technique for eliciting reasoning from Large Language Models (LLMs), yet the ro

model-releasesarxiv-cs-ai
20 Apr 2026
Applications

Information Router for Mitigating Modality Dominance in Vision-Language Models

DGX agent

arXiv:2604.16264v1 Announce Type: new Abstract: Vision Language models (VLMs) have demonstrated strong performance across a wide range of benchmarks, yet they often suffer from modality dominance, whe

applicationsarxiv-cs-cv
20 Apr 2026
Safety

Language Models as Semantic Teachers: Post-Training Alignment for Medical Audio Understanding

DGX agent

arXiv:2512.04847v2 Announce Type: replace-cross Abstract: Pre-trained audio models excel at detecting acoustic patterns in auscultation sounds but often fail to grasp their clinical significance, limi

safetyarxiv-cs-ai
20 Apr 2026
Research

Modeling Parkinson's Disease Progression Using Longitudinal Voice Biomarkers: A Comparative Study of Statistical and Neural Mixed-Effects Models

DGX agent

arXiv:2507.20058v3 Announce Type: replace-cross Abstract: Predicting Parkinson's Disease (PD) progression is crucial for personalized treatment, and voice biomarkers offer a promising non-invasive met

researcharxiv-cs-lg
20 Apr 2026
Research

DySCO: Dynamic Attention-Scaling Decoding for Long-Context Language Models

DGX agent

arXiv:2602.22175v2 Announce Type: replace Abstract: Understanding and reasoning over long contexts is a crucial capability for language models (LMs). Although recent models support increasingly long c

researcharxiv-cs-cl
17 Apr 2026
Model Releases

IF-RewardBench: Benchmarking Judge Models for Instruction-Following Evaluation

DGX agent

arXiv:2603.04738v2 Announce Type: replace Abstract: Instruction-following is a foundational capability of large language models (LLMs), with its improvement hinging on scalable and accurate feedback f

model-releasesarxiv-cs-cl
17 Apr 2026
Safety

Learning Ad Hoc Network Dynamics via Graph-Structured World Models

DGX agent

arXiv:2604.14811v1 Announce Type: new Abstract: Ad hoc wireless networks exhibit complex, innate and coupled dynamics: node mobility, energy depletion and topology change that are difficult to model a

safetyarxiv-cs-lg
17 Apr 2026
Model Releases

Beyond Static Personas: Situational Personality Steering for Large Language Models

DGX agent

arXiv:2604.13846v1 Announce Type: new Abstract: Personalized Large Language Models (LLMs) facilitate more natural, human-like interactions in human-centric applications. However, existing personalizat

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Peer-Predictive Self-Training for Language Model Reasoning

DGX agent

arXiv:2604.13356v1 Announce Type: new Abstract: Mechanisms for continued self-improvement of language models without external supervision remain an open challenge. We propose Peer-Predictive Self-Trai

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

The Consciousness Cluster: Emergent preferences of Models that Claim to be Conscious

DGX agent

arXiv:2604.13051v1 Announce Type: new Abstract: There is debate about whether LLMs can be conscious. We investigate a distinct question: if a model claims to be conscious, how does this affect its dow

model-releasesarxiv-cs-cl
16 Apr 2026
Research

Cross-Domain Transfer with Particle Physics Foundation Models: From Jets to Neutrino Interactions

DGX agent

arXiv:2604.12364v1 Announce Type: cross Abstract: Future AI-based studies in particle physics will likely start from a foundation model to accelerate training and enhance sensitivity. As a step toward

researcharxiv-cs-lg
15 Apr 2026
Model Releases

FABLE: Fine-grained Fact Anchoring for Unstructured Model Editing

DGX agent

arXiv:2604.12559v1 Announce Type: new Abstract: Unstructured model editing aims to update models with real-world text, yet existing methods often memorize text holistically without reliable fine-grain

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

Filtered Reasoning Score: Evaluating Reasoning Quality on a Model's Most-Confident Traces

DGX agent

arXiv:2604.11996v1 Announce Type: cross Abstract: Should we trust Large Language Models (LLMs) with high accuracy? LLMs achieve high accuracy on reasoning benchmarks, but correctness alone does not re

model-releasesarxiv-cs-ai
15 Apr 2026
Research

LoSA: Locality Aware Sparse Attention for Block-Wise Diffusion Language Models

DGX agent

arXiv:2604.12056v1 Announce Type: new Abstract: Block-wise diffusion language models (DLMs) generate multiple tokens in any order, offering a promising alternative to the autoregressive decoding pipel

researcharxiv-cs-cl
15 Apr 2026
Model Releases

MoshiRAG: Asynchronous Knowledge Retrieval for Full-Duplex Speech Language Models

DGX agent

arXiv:2604.12928v1 Announce Type: new Abstract: Speech-to-speech language models have recently emerged to enhance the naturalness of conversational AI. In particular, full-duplex models are distinguis

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

Towards EnergyGPT: A Large Language Model Specialized for the Energy Sector

DGX agent

arXiv:2509.07177v3 Announce Type: replace Abstract: Large language models have demonstrated impressive capabilities across various domains. However, their general-purpose nature often limits their eff

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

Benchmarking Vision-Language Models under Contradictory Virtual Content Attacks in Augmented Reality

DGX agent

arXiv:2604.05510v2 Announce Type: replace Abstract: Augmented reality (AR) has rapidly expanded over the past decade. As AR becomes increasingly integrated into daily life, its security and reliabilit

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Consistency of AI-Generated Exercise Prescriptions: A Repeated Generation Study Using a Large Language Model

DGX agent

arXiv:2604.11287v1 Announce Type: new Abstract: Background: Large language models (LLMs) have been explored as tools for generating personalized exercise prescriptions, yet the consistency of outputs

model-releasesarxiv-cs-ai
14 Apr 2026
Tutorials

Continuous Adversarial Flow Models

DGX agent

arXiv:2604.11521v1 Announce Type: cross Abstract: We propose continuous adversarial flow models, a type of continuous-time flow model trained with an adversarial objective. Unlike flow matching, which

tutorialsarxiv-cs-cv
14 Apr 2026
Model Releases

FS-DFM: Fast and Accurate Long Text Generation with Few-Step Diffusion Language Models

DGX agent

arXiv:2509.20624v5 Announce Type: replace-cross Abstract: Autoregressive language models (ARMs) deliver strong likelihoods, but are inherently serial: they generate one token per forward pass, which l

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Grid2Matrix: Revealing Digital Agnosia in Vision-Language Models

DGX agent

arXiv:2604.09687v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) excel on many multimodal reasoning benchmarks, but these evaluations often do not require an exhaustive readout of the i

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

How Many Tries Does It Take? Iterative Self-Repair in LLM Code Generation Across Model Scales and Benchmarks

DGX agent

arXiv:2604.10508v1 Announce Type: cross Abstract: Large language models frequently fail to produce correct code on their first attempt, yet most benchmarks evaluate them in a single-shot setting. We i

model-releasesarxiv-cs-ai
14 Apr 2026
Applications

Modular Delta Merging with Orthogonal Constraints: A Scalable Framework for Continual and Reversible Model Composition

DGX agent

arXiv:2507.20997v4 Announce Type: replace-cross Abstract: In real-world machine learning deployments, models must be continually updated, composed, and when required, selectively undone. However, exis

applicationsarxiv-cs-ai
14 Apr 2026
Research

Parallelism and Generation Order in Masked Diffusion Language Models: Limits Today, Potential Tomorrow

DGX agent

arXiv:2601.15593v2 Announce Type: replace-cross Abstract: Masked Diffusion Language Models (MDLMs) promise parallel token generation and arbitrary-order decoding, yet it remains unclear to what extent

researcharxiv-cs-ai
14 Apr 2026
Model Releases

Self-supervised Pretraining of Cell Segmentation Models

DGX agent

arXiv:2604.10609v1 Announce Type: new Abstract: Instance segmentation enables the analysis of spatial and temporal properties of cells in microscopy images by identifying the pixels belonging to each

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

SRBench: A Comprehensive Benchmark for Sequential Recommendation with Large Language Models

DGX agent

arXiv:2604.09553v1 Announce Type: cross Abstract: LLM development has aroused great interest in Sequential Recommendation (SR) applications. However, comprehensive evaluation of SR models remains lack

model-releasesarxiv-cs-ai
14 Apr 2026
Tutorials

Symmetry-Aware Generative Modeling through Learned Canonicalization

DGX agent

arXiv:2501.07773v3 Announce Type: replace Abstract: Generative modeling of symmetric densities has a range of applications in AI for science, from drug discovery to physics simulations. The existing g

tutorialsarxiv-cs-lg
14 Apr 2026
Model Releases

TorchUMM: A Unified Multimodal Model Codebase for Evaluation, Analysis, and Post-training

DGX agent

arXiv:2604.10784v1 Announce Type: new Abstract: Recent advances in unified multimodal models (UMMs) have led to a proliferation of architectures capable of understanding, generating, and editing acros

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Distilling Genomic Models for Efficient mRNA Representation Learning via Embedding Matching

DGX agent

arXiv:2604.08574v1 Announce Type: cross Abstract: Large Genomic Foundation Models have recently achieved remarkable results and in-vivo translation capabilities. However these models quickly grow to o

researcharxiv-cs-ai
13 Apr 2026
Safety

Learning Vision-Language-Action World Models for Autonomous Driving

DGX agent

arXiv:2604.09059v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have recently achieved notable progress in end-to-end autonomous driving by integrating perception, reasoning, and

safetyarxiv-cs-ai
13 Apr 2026
Safety

Post-Selection Distributional Model Evaluation

DGX agent

arXiv:2603.23055v2 Announce Type: replace-cross Abstract: Formal model evaluation methods typically certify that a model satisfies a prescribed target key performance indicator (KPI) level. However, i

safetyarxiv-cs-lg
13 Apr 2026
Model Releases

UAV-Track VLA: Embodied Aerial Tracking via Vision-Language-Action Models

DGX agent

arXiv:2604.02241v2 Announce Type: replace Abstract: Embodied visual tracking is crucial for Unmanned Aerial Vehicles (UAVs) executing complex real-world tasks. In dynamic urban scenarios with complex

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

Ads in AI Chatbots? An Analysis of How Large Language Models Navigate Conflicts of Interest

DGX agent

arXiv:2604.08525v1 Announce Type: cross Abstract: Today's large language models (LLMs) are trained to align with user preferences through methods such as reinforcement learning. Yet models are beginni

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Beyond Social Pressure: Benchmarking Epistemic Attack in Large Language Models

DGX agent

arXiv:2604.07749v1 Announce Type: new Abstract: Large language models (LLMs) can shift their answers under pressure in ways that reflect accommodation rather than reasoning. Prior work on sycophancy h

model-releasesarxiv-cs-cl
10 Apr 2026
← Previous
1…2021222324…1012
Next →