AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

88,271Total entries
1Added by human
88,270Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,519 results
24 Apr 2026

climt-paraformer: Stable Emulation of Convective Parameterization using a Temporal Memory-aware Transformer

ResearchDGX agent

arXiv:2604.21085v1 Announce Type: cross Abstract: Accurate representation of moist convective sub-grid-scale processes remains a major challenge in global climate models, as traditional parameterizati

CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors

Model ReleasesDGX agent

arXiv:2604.21241v1 Announce Type: cross Abstract: Vision--Language--Action (VLA) models often use intermediate representations to connect multimodal inputs with continuous control, yet spatial guidanc

Data-Driven Open-Loop Simulation for Digital-Twin Operator Decision Support in Wastewater Treatment

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.20935v1 Announce Type: cross Abstract: Wastewater treatment plants (WWTPs) need digital-twin-style decision support tools that can simulate plant response under prescribed control plans, to

DeepSeek v4 just dropped

Model ReleasesDGX agent

DeepSeek has released v4, its latest model iteration. The announcement was made by Clem Delangue on X (formerly Twitter). This likely represents a significant update to DeepSeek's AI capabilities, tho

DeepSeek V4 Pro is now available on Together AI. DeepSeek V4 Flash coming soon. Try it now: http://www.together.ai/models/deepseek-v4-pro#

Model ReleasesDGX agent

DeepSeek V4 Pro is now available through Together AI's model platform, with the faster DeepSeek V4 Flash variant expected to launch soon. Together AI is offering users the ability to access and test D

Dialect vs Demographics: Quantifying LLM Bias from Implicit Linguistic Signals vs. Explicit User Profiles

Model ReleasesDGX agent

arXiv:2604.21152v1 Announce Type: cross Abstract: As state-of-the-art Large Language Models (LLMs) have become ubiquitous, ensuring equitable performance across diverse demographics is critical. Howev

Do MLLMs Understand Pointing? Benchmarking and Enhancing Referential Reasoning in Egocentric Vision

Model ReleasesDGX agent

arXiv:2604.21461v1 Announce Type: new Abstract: Egocentric AI agents, such as smart glasses, rely on pointing gestures to resolve referential ambiguities in natural language commands. However, despite

FairyFuse: Multiplication-Free LLM Inference on CPUs via Fused Ternary Kernels

Model ReleasesDGX agent

arXiv:2604.20913v1 Announce Type: new Abstract: Large language models are increasingly deployed on CPU-only platforms where memory bandwidth is the primary bottleneck for autoregressive generation. We

Fine-Tuning Regimes Define Distinct Continual Learning Problems

Model ReleasesDGX agent

arXiv:2604.21927v1 Announce Type: new Abstract: Continual learning (CL) studies how models acquire tasks sequentially while retaining previously learned knowledge. Despite substantial progress in benc

Flipping Against All Odds: Reducing LLM Coin Flip Bias via Verbalized Rejection Sampling

SafetyDGX agent

arXiv:2506.09998v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can often accurately describe probability distributions using natural language, yet they still struggle to genera

Generative Discovery of Magnetic Insulators under Competing Physical Constraints

ResearchDGX agent

arXiv:2604.21073v1 Announce Type: cross Abstract: Discovering materials that must simultaneously satisfy multiple competing constraints remains a central challenge in computational materials design, p

GerAV: Towards New Heights in German Authorship Verification using Fine-Tuned LLMs on a New Benchmark

Model ReleasesDGX agent

arXiv:2601.13711v2 Announce Type: replace Abstract: Authorship verification (AV) is the task of determining whether two texts were written by the same author and has been studied extensively, predomin

GiVA: Gradient-Informed Bases for Vector-Based Adaptation

Model ReleasesDGX agent

arXiv:2604.21901v1 Announce Type: cross Abstract: As model sizes continue to grow, parameter-efficient fine-tuning has emerged as a powerful alternative to full fine-tuning. While LoRA is widely adopt

GPT-5.5 now available in Deep Agents!

Model ReleasesDGX agent

GPT-5.5 now available in Deep Agents! GPT-5.5 is now available in the API. The model brings higher intelligence and stronger token efficiency to complex work, helping tasks get done with fewer retries

HARBOR: Automated Harness Optimization

SafetyDGX agent

arXiv:2604.20938v1 Announce Type: cross Abstract: Long-horizon language-model agents are dominated, in lines of code and in operational complexity, not by their underlying model but by the harness tha

Here's DeepSeek v4 Pro. Added to the playable gallery as well.

Model ReleasesDGX agent

Here's DeepSeek v4 Pro. Added to the playable gallery as well. Media I had a range of models 'build me a procedurally generated 3D simulation showing the evolution of a harbor town from 3000 BCE to 30

Information Bottleneck-Guided Heterogeneous Graph Learning for Interpretable Neurodevelopmental Disorder Diagnosis

TutorialsDGX agent

arXiv:2502.20769v3 Announce Type: replace Abstract: Developing interpretable models for neurodevelopmental disorders (NDDs) diagnosis presents significant challenges in effectively encoding, decoding,

Interpretable facial dynamics as behavioral and perceptual traces of deepfakes

Model ReleasesDGX agent

arXiv:2604.21760v1 Announce Type: new Abstract: Deepfake detection research has largely converged on deep learning approaches that, despite strong benchmark performance, offer limited insight into wha

Job Skill Extraction via LLM-Centric Multi-Module Framework

ResearchDGX agent

arXiv:2604.21525v1 Announce Type: new Abstract: Span-level skill extraction from job advertisements underpins candidate-job matching and labor-market analytics, yet generative large language models (L

LAF-Based Evaluation and UTTL-Based Learning Strategies with MIATTs

ApplicationsDGX agent

arXiv:2604.20944v1 Announce Type: cross Abstract: In many real-world machine learning (ML) applications, the true target cannot be precisely defined due to ambiguity or subjectivity information. To ad

LatRef-Diff: Latent and Reference-Guided Diffusion for Facial Attribute Editing and Style Manipulation

ResearchDGX agent

arXiv:2604.21279v1 Announce Type: new Abstract: Facial attribute editing and style manipulation are crucial for applications like virtual avatars and photo editing. However, achieving precise control

Listen and Chant Before You Read: The Ladder of Beauty in LM Pre-Training

ResearchDGX agent

arXiv:2604.21265v1 Announce Type: new Abstract: We show that pre-training a Transformer on music before language significantly accelerates language acquisition. Using piano performances (MAESTRO datas

Logic Jailbreak: Efficiently Unlocking LLM Safety Restrictions Through Formal Logical Expression

SafetyDGX agent

arXiv:2505.13527v3 Announce Type: replace-cross Abstract: Despite substantial advancements in aligning large language models (LLMs) with human values, current safety mechanisms remain susceptible to j

Losing our Tail, Again: (Un)Natural Selection & Multilingual LLMs

ResearchDGX agent

arXiv:2507.03933v3 Announce Type: replace Abstract: Multilingual Large Language Models considerably changed how technologies influence language. While previous technologies could mediate or assist hum

MaskDiME: Adaptive Masked Diffusion for Precise and Efficient Visual Counterfactual Explanations

Model ReleasesDGX agent

arXiv:2602.18792v3 Announce Type: replace Abstract: Visual counterfactual explanations aim to reveal the minimal semantic modifications that can alter a model's prediction, providing causal and interp

Mind the Prompt: Self-adaptive Generation of Task Plan Explanations via LLMs

SafetyDGX agent

arXiv:2604.21092v1 Announce Type: new Abstract: Integrating Large Language Models (LLMs) into complex software systems enables the generation of human-understandable explanations of opaque AI processe

Mixture of Sequence: Theme-Aware Mixture-of-Experts for Long-Sequence Recommendation

TutorialsDGX agent

arXiv:2604.20858v1 Announce Type: cross Abstract: Sequential recommendation has rapidly advanced in click-through rate prediction due to its ability to model dynamic user interests. A key challenge, h

Omission Constraints Decay While Commission Constraints Persist in Long-Context LLM Agents

Model ReleasesDGX agent

arXiv:2604.20911v1 Announce Type: cross Abstract: LLM agents deployed in production operate under operator-defined behavioral policies (system-prompt instructions such as prohibitions on credential di

OmniFit: Multi-modal 3D Body Fitting via Scale-agnostic Dense Landmark Prediction

Model ReleasesDGX agent

arXiv:2604.21575v1 Announce Type: new Abstract: Fitting an underlying body model to 3D clothed human assets has been extensively studied, yet most approaches focus on either single-modal inputs such a

Optimizing Diffusion Priors with a Single Observation

ApplicationsDGX agent

arXiv:2604.21066v1 Announce Type: new Abstract: While diffusion priors generate high-quality posterior samples across many inverse problems, they are often trained on limited training sets or purely s

Our teams have been busyyy! Here are some key updates from the past week: — @GoogleCloud unveiled a suite of AI innovations at our Cloud Nex…

Model ReleasesDGX agent

Our teams have been busyyy! Here are some key updates from the past week: — @GoogleCloud unveiled a suite of AI innovations at our Cloud Next event, including our eighth generation TPUs (TPUt for infe

PAT3D: Physics-Augmented Text-to-3D Scene Generation

ResearchDGX agent

arXiv:2511.21978v2 Announce Type: replace Abstract: We introduce PAT3D, the first physics-augmented text-to-3D scene generation framework that integrates vision-language models with physics-based simu

Pre-trained LLMs Meet Sequential Recommenders: Efficient User-Centric Knowledge Distillation

ResearchDGX agent

arXiv:2604.21536v1 Announce Type: cross Abstract: Sequential recommender systems have achieved significant success in modeling temporal user behavior but remain limited in capturing rich user semantic

Preferences of a Voice-First Nation: Large-Scale Pairwise Evaluation and Preference Analysis for TTS in Indian Languages

ResearchDGX agent

arXiv:2604.21481v1 Announce Type: new Abstract: Crowdsourced pairwise evaluation has emerged as a scalable approach for assessing foundation models. However, applying it to Text to Speech(TTS) introdu

Propensity Inference: Environmental Contributors to LLM Behaviour

ResearchDGX agent

arXiv:2604.21098v1 Announce Type: new Abstract: Motivated by loss of control risks from misaligned AI systems, we develop and apply methods for measuring language models' propensity for unsanctioned b

Reasoning About Traversability: Language-Guided Off-Road 3D Trajectory Planning

Model ReleasesDGX agent

arXiv:2604.21249v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) enable high-level semantic reasoning for end-to-end autonomous driving, particularly in unstructured environments, e

Reinforcing privacy reasoning in LLMs via normative simulacra from fiction

Model ReleasesDGX agent

arXiv:2604.20904v1 Announce Type: cross Abstract: Information handling practices of LLM agents are broadly misaligned with the contextual privacy expectations of their users. Contextual Integrity (CI)

Remember o3 was only a year and a week ago! Also, only GPT-5.5 seemed to take the 'evolution' piece seriously and change the setting rather …

Model ReleasesDGX agent

I cannot provide an accurate summary for this entry as the text appears incomplete and lacks sufficient context. The post fragment references o3 (likely an AI model), GPT-5.5, and discusses timeline/e

Sink-Token-Aware Pruning for Fine-Grained Video Understanding in Efficient Video LLMs

ResearchDGX agent

arXiv:2604.20937v1 Announce Type: new Abstract: Video Large Language Models (Video LLMs) incur high inference latency due to a large number of visual tokens provided to LLMs. To address this, training

SparseGF: A Height-Aware Sparse Segmentation Framework with Context Compression for Robust Ground Filtering Across Urban to Natural Scenes

Model ReleasesDGX agent

arXiv:2604.21356v1 Announce Type: new Abstract: High-quality digital terrain models derived from airborne laser scanning (ALS) data are essential for a wide range of geospatial analyses, and their gen

Strategic Heterogeneous Multi-Agent Architecture for Cost-Effective Code Vulnerability Detection

Model ReleasesDGX agent

arXiv:2604.21282v1 Announce Type: cross Abstract: Automated code vulnerability detection is critical for software security, yet existing approaches face a fundamental trade-off between detection accur

The Coding Assistant Breakdown: More Tokens Please

Model ReleasesDGX agent

This analysis examines the token consumption and economics of coding assistants, likely comparing different AI models' efficiency and cost-effectiveness for code generation tasks. The piece probably d

The Root Theorem of Context Engineering

Model ReleasesDGX agent

arXiv:2604.20874v1 Announce Type: cross Abstract: Every system that maintains a large language model conversation beyond a single session faces two inescapable constraints: the context window is finit

TRACES: Tagging Reasoning Steps for Adaptive Cost-Efficient Early-Stopping

ResearchDGX agent

arXiv:2604.21057v1 Announce Type: new Abstract: The field of Language Reasoning Models (LRMs) has been very active over the past few years with advances in training and inference techniques enabling L

Transferable SCF-Acceleration through Solver-Aligned Initialization Learning

ResearchDGX agent

arXiv:2604.21657v1 Announce Type: new Abstract: Machine learning methods that predict initial guesses from molecular geometry can reduce this cost, but matrix-prediction models fail when extrapolating

VVS: Accelerating Speculative Decoding for Visual Autoregressive Generation via Partial Verification Skipping

ResearchDGX agent

arXiv:2511.13587v2 Announce Type: replace-cross Abstract: Visual autoregressive (AR) generation models have demonstrated strong potential for image generation, yet their next-token-prediction paradigm

We benchmarked GPT-5.5 on document understanding 📄📊 We ran it through ParseBench, our comprehensive OCR benchmark over enterprise document…

Model ReleasesDGX agent

We benchmarked GPT-5.5 on document understanding 📄📊 We ran it through ParseBench, our comprehensive OCR benchmark over enterprise documents. We evaluated metrics across various dimensions: visual grou

When to Trust the Answer: Question-Aligned Semantic Nearest Neighbor Entropy for Safer Surgical VQA

Model ReleasesDGX agent

arXiv:2511.01458v2 Announce Type: replace-cross Abstract: Safety and reliability are critical for deploying visual question answering (VQA) systems in surgery, where incorrect or ambiguous responses c

XtraGPT: Context-Aware and Controllable Academic Paper Revision via Human-AI Collaboration

SafetyDGX agent

arXiv:2505.11336v4 Announce Type: replace Abstract: Despite the growing adoption of large language models (LLMs) in academic workflows, their capabilities remain limited in supporting high-quality sci

23 Apr 2026

Analyzing Shapley Additive Explanations to Understand Anomaly Detection Algorithm Behaviors and Their Complementarity

ResearchDGX agent

arXiv:2602.00208v2 Announce Type: replace-cross Abstract: Unsupervised anomaly detection is a challenging problem due to the diversity of data distributions and the lack of labels. Ensemble methods ar

API pricing will be 5 per 1 million input tokens and 30 per 1 million output tokens, with a 1 million context window. (Remember, you will …

Model ReleasesDGX agent

OpenAI's API pricing structure charges 5 per 1 million input tokens and 30 per 1 million output tokens, with support for a 1 million token context window. This pricing model reflects the higher cost o

Automatic Ontology Construction Using LLMs as an External Layer of Memory, Verification, and Planning for Hybrid Intelligent Systems

Model ReleasesDGX agent

arXiv:2604.20795v1 Announce Type: new Abstract: This paper presents a hybrid architecture for intelligent systems in which large language models (LLMs) are extended with an external ontological memory

Beyond Majority Voting: Towards Fine-grained and More Reliable Reward Signal for Test-Time Reinforcement Learning

Model ReleasesDGX agent

arXiv:2512.15146v3 Announce Type: replace Abstract: Test-time reinforcement learning mitigates the reliance on annotated data by using majority voting results as pseudo-labels, emerging as a complemen

Bias in the Tails: How Name-conditioned Evaluative Framing in Resume Summaries Destabilizes LLM-based Hiring

SafetyDGX agent

arXiv:2604.19984v1 Announce Type: cross Abstract: Research has documented LLMs' name-based bias in hiring and salary recommendations. In this paper, we instead consider a setting where LLMs generate c

Bimanual Robot Manipulation via Multi-Agent In-Context Learning

Model ReleasesDGX agent

arXiv:2604.20348v1 Announce Type: cross Abstract: Language Models (LLMs) have emerged as powerful reasoning engines for embodied control. In particular, In-Context Learning (ICL) enables off-the-shelf

Coding with Eyes: Visual Feedback Unlocks Reliable GUI Code Generating and Debugging

Model ReleasesDGX agent

arXiv:2604.19750v1 Announce Type: cross Abstract: Recent advances in Large Language Model (LLM)-based agents have shown remarkable progress in code generation. However, current agent methods mainly re

Cognitive Alignment At No Cost: Inducing Human Attention Biases For Interpretable Vision Transformers

SafetyDGX agent

arXiv:2604.20027v1 Announce Type: cross Abstract: For state-of-the-art image understanding, Vision Transformers (ViTs) have become the standard architecture but their processing diverges substantially

Continuous Semantic Caching for Low-Cost LLM Serving

ApplicationsDGX agent

arXiv:2604.20021v1 Announce Type: cross Abstract: As Large Language Models (LLMs) become increasingly popular, caching responses so that they can be reused by users with semantically similar queries h

Differentiable Conformal Training for LLM Reasoning Factuality

Model ReleasesDGX agent

arXiv:2604.20098v1 Announce Type: new Abstract: Large Language Models (LLMs) frequently hallucinate, limiting their reliability in critical applications. Conformal Prediction (CP) addresses this by ca

Epistemic Constitutionalism Or: how to avoid coherence bias

SafetyDGX agent

arXiv:2601.14295v3 Announce Type: replace Abstract: Large language models increasingly function as artificial reasoners: they evaluate arguments, assign credibility, and express confidence. Yet their

← Previous
1…424425426427428…1059
Next →