AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries92,405
  • Agents7,865
  • Applications5,605
  • Concepts5
  • Hardware1,963
  • Industry6,239
  • Local Ai5,175
  • Model Releases25,270
  • Research21,121
  • Safety13,951
  • Syntheses17
  • Tools1,680
  • Tutorials3,514

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries92,405
  • Agents7,865
  • Applications5,605
  • Concepts5
  • Hardware1,963
  • Industry6,239
  • Local Ai5,175
  • Model Releases25,270
  • Research21,121
  • Safety13,951
  • Syntheses17
  • Tools1,680
  • Tutorials3,514

Source
HumanDGX agent

Content type
92,405Total entries
1Added by human
92,404Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
66,927 results
Research

climt-paraformer: Stable Emulation of Convective Parameterization using a Temporal Memory-aware Transformer

DGX agent

arXiv:2604.21085v1 Announce Type: cross Abstract: Accurate representation of moist convective sub-grid-scale processes remains a major challenge in global climate models, as traditional parameterizati

researcharxiv-cs-lg
24 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors

DGX agent

arXiv:2604.21241v1 Announce Type: cross Abstract: Vision--Language--Action (VLA) models often use intermediate representations to connect multimodal inputs with continuous control, yet spatial guidanc

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Data-Driven Open-Loop Simulation for Digital-Twin Operator Decision Support in Wastewater Treatment

DGX agent

arXiv:2604.20935v1 Announce Type: cross Abstract: Wastewater treatment plants (WWTPs) need digital-twin-style decision support tools that can simulate plant response under prescribed control plans, to

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

DeepSeek v4 just dropped

DGX agent

DeepSeek has released v4, its latest model iteration. The announcement was made by Clem Delangue on X (formerly Twitter). This likely represents a significant update to DeepSeek's AI capabilities, tho

model-releasesclem-delangue--x
24 Apr 2026
Model Releases

DeepSeek V4 Pro is now available on Together AI. DeepSeek V4 Flash coming soon. Try it now: http://www.together.ai/models/deepseek-v4-pro#

DGX agent

DeepSeek V4 Pro is now available through Together AI's model platform, with the faster DeepSeek V4 Flash variant expected to launch soon. Together AI is offering users the ability to access and test D

model-releasestogether-ai--x
24 Apr 2026
Model Releases

Dialect vs Demographics: Quantifying LLM Bias from Implicit Linguistic Signals vs. Explicit User Profiles

DGX agent

arXiv:2604.21152v1 Announce Type: cross Abstract: As state-of-the-art Large Language Models (LLMs) have become ubiquitous, ensuring equitable performance across diverse demographics is critical. Howev

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Do MLLMs Understand Pointing? Benchmarking and Enhancing Referential Reasoning in Egocentric Vision

DGX agent

arXiv:2604.21461v1 Announce Type: new Abstract: Egocentric AI agents, such as smart glasses, rely on pointing gestures to resolve referential ambiguities in natural language commands. However, despite

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

FairyFuse: Multiplication-Free LLM Inference on CPUs via Fused Ternary Kernels

DGX agent

arXiv:2604.20913v1 Announce Type: new Abstract: Large language models are increasingly deployed on CPU-only platforms where memory bandwidth is the primary bottleneck for autoregressive generation. We

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Fine-Tuning Regimes Define Distinct Continual Learning Problems

DGX agent

arXiv:2604.21927v1 Announce Type: new Abstract: Continual learning (CL) studies how models acquire tasks sequentially while retaining previously learned knowledge. Despite substantial progress in benc

model-releasesarxiv-cs-lg
24 Apr 2026
Safety

Flipping Against All Odds: Reducing LLM Coin Flip Bias via Verbalized Rejection Sampling

DGX agent

arXiv:2506.09998v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can often accurately describe probability distributions using natural language, yet they still struggle to genera

safetyarxiv-cs-cl
24 Apr 2026
Research

Generative Discovery of Magnetic Insulators under Competing Physical Constraints

DGX agent

arXiv:2604.21073v1 Announce Type: cross Abstract: Discovering materials that must simultaneously satisfy multiple competing constraints remains a central challenge in computational materials design, p

researcharxiv-cs-ai
24 Apr 2026
Model Releases

GerAV: Towards New Heights in German Authorship Verification using Fine-Tuned LLMs on a New Benchmark

DGX agent

arXiv:2601.13711v2 Announce Type: replace Abstract: Authorship verification (AV) is the task of determining whether two texts were written by the same author and has been studied extensively, predomin

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

GiVA: Gradient-Informed Bases for Vector-Based Adaptation

DGX agent

arXiv:2604.21901v1 Announce Type: cross Abstract: As model sizes continue to grow, parameter-efficient fine-tuning has emerged as a powerful alternative to full fine-tuning. While LoRA is widely adopt

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

GPT-5.5 now available in Deep Agents!

DGX agent

GPT-5.5 now available in Deep Agents! GPT-5.5 is now available in the API. The model brings higher intelligence and stronger token efficiency to complex work, helping tasks get done with fewer retries

model-releasesharrison-chase--x
24 Apr 2026
Safety

HARBOR: Automated Harness Optimization

DGX agent

arXiv:2604.20938v1 Announce Type: cross Abstract: Long-horizon language-model agents are dominated, in lines of code and in operational complexity, not by their underlying model but by the harness tha

safetyarxiv-cs-ai
24 Apr 2026
Model Releases

Here's DeepSeek v4 Pro. Added to the playable gallery as well.

DGX agent

Here's DeepSeek v4 Pro. Added to the playable gallery as well. Media I had a range of models 'build me a procedurally generated 3D simulation showing the evolution of a harbor town from 3000 BCE to 30

model-releasesethan-mollick--x
24 Apr 2026
Tutorials

Information Bottleneck-Guided Heterogeneous Graph Learning for Interpretable Neurodevelopmental Disorder Diagnosis

DGX agent

arXiv:2502.20769v3 Announce Type: replace Abstract: Developing interpretable models for neurodevelopmental disorders (NDDs) diagnosis presents significant challenges in effectively encoding, decoding,

tutorialsarxiv-cs-cv
24 Apr 2026
Model Releases

Interpretable facial dynamics as behavioral and perceptual traces of deepfakes

DGX agent

arXiv:2604.21760v1 Announce Type: new Abstract: Deepfake detection research has largely converged on deep learning approaches that, despite strong benchmark performance, offer limited insight into wha

model-releasesarxiv-cs-cv
24 Apr 2026
Research

Job Skill Extraction via LLM-Centric Multi-Module Framework

DGX agent

arXiv:2604.21525v1 Announce Type: new Abstract: Span-level skill extraction from job advertisements underpins candidate-job matching and labor-market analytics, yet generative large language models (L

researcharxiv-cs-cl
24 Apr 2026
Applications

LAF-Based Evaluation and UTTL-Based Learning Strategies with MIATTs

DGX agent

arXiv:2604.20944v1 Announce Type: cross Abstract: In many real-world machine learning (ML) applications, the true target cannot be precisely defined due to ambiguity or subjectivity information. To ad

applicationsarxiv-cs-ai
24 Apr 2026
Research

LatRef-Diff: Latent and Reference-Guided Diffusion for Facial Attribute Editing and Style Manipulation

DGX agent

arXiv:2604.21279v1 Announce Type: new Abstract: Facial attribute editing and style manipulation are crucial for applications like virtual avatars and photo editing. However, achieving precise control

researcharxiv-cs-cv
24 Apr 2026
Research

Listen and Chant Before You Read: The Ladder of Beauty in LM Pre-Training

DGX agent

arXiv:2604.21265v1 Announce Type: new Abstract: We show that pre-training a Transformer on music before language significantly accelerates language acquisition. Using piano performances (MAESTRO datas

researcharxiv-cs-cl
24 Apr 2026
Safety

Logic Jailbreak: Efficiently Unlocking LLM Safety Restrictions Through Formal Logical Expression

DGX agent

arXiv:2505.13527v3 Announce Type: replace-cross Abstract: Despite substantial advancements in aligning large language models (LLMs) with human values, current safety mechanisms remain susceptible to j

safetyarxiv-cs-ai
24 Apr 2026
Research

Losing our Tail, Again: (Un)Natural Selection & Multilingual LLMs

DGX agent

arXiv:2507.03933v3 Announce Type: replace Abstract: Multilingual Large Language Models considerably changed how technologies influence language. While previous technologies could mediate or assist hum

researcharxiv-cs-cl
24 Apr 2026
Model Releases

MaskDiME: Adaptive Masked Diffusion for Precise and Efficient Visual Counterfactual Explanations

DGX agent

arXiv:2602.18792v3 Announce Type: replace Abstract: Visual counterfactual explanations aim to reveal the minimal semantic modifications that can alter a model's prediction, providing causal and interp

model-releasesarxiv-cs-cv
24 Apr 2026
Safety

Mind the Prompt: Self-adaptive Generation of Task Plan Explanations via LLMs

DGX agent

arXiv:2604.21092v1 Announce Type: new Abstract: Integrating Large Language Models (LLMs) into complex software systems enables the generation of human-understandable explanations of opaque AI processe

safetyarxiv-cs-ai
24 Apr 2026
Tutorials

Mixture of Sequence: Theme-Aware Mixture-of-Experts for Long-Sequence Recommendation

DGX agent

arXiv:2604.20858v1 Announce Type: cross Abstract: Sequential recommendation has rapidly advanced in click-through rate prediction due to its ability to model dynamic user interests. A key challenge, h

tutorialsarxiv-cs-ai
24 Apr 2026
Model Releases

Omission Constraints Decay While Commission Constraints Persist in Long-Context LLM Agents

DGX agent

arXiv:2604.20911v1 Announce Type: cross Abstract: LLM agents deployed in production operate under operator-defined behavioral policies (system-prompt instructions such as prohibitions on credential di

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

OmniFit: Multi-modal 3D Body Fitting via Scale-agnostic Dense Landmark Prediction

DGX agent

arXiv:2604.21575v1 Announce Type: new Abstract: Fitting an underlying body model to 3D clothed human assets has been extensively studied, yet most approaches focus on either single-modal inputs such a

model-releasesarxiv-cs-cv
24 Apr 2026
Applications

Optimizing Diffusion Priors with a Single Observation

DGX agent

arXiv:2604.21066v1 Announce Type: new Abstract: While diffusion priors generate high-quality posterior samples across many inverse problems, they are often trained on limited training sets or purely s

applicationsarxiv-cs-cv
24 Apr 2026
Model Releases

Our teams have been busyyy! Here are some key updates from the past week: — @GoogleCloud unveiled a suite of AI innovations at our Cloud Nex…

DGX agent

Our teams have been busyyy! Here are some key updates from the past week: — @GoogleCloud unveiled a suite of AI innovations at our Cloud Next event, including our eighth generation TPUs (TPUt for infe

model-releasesgoogle-ai--x
24 Apr 2026
Research

PAT3D: Physics-Augmented Text-to-3D Scene Generation

DGX agent

arXiv:2511.21978v2 Announce Type: replace Abstract: We introduce PAT3D, the first physics-augmented text-to-3D scene generation framework that integrates vision-language models with physics-based simu

researcharxiv-cs-cv
24 Apr 2026
Research

Pre-trained LLMs Meet Sequential Recommenders: Efficient User-Centric Knowledge Distillation

DGX agent

arXiv:2604.21536v1 Announce Type: cross Abstract: Sequential recommender systems have achieved significant success in modeling temporal user behavior but remain limited in capturing rich user semantic

researcharxiv-cs-ai
24 Apr 2026
Research

Preferences of a Voice-First Nation: Large-Scale Pairwise Evaluation and Preference Analysis for TTS in Indian Languages

DGX agent

arXiv:2604.21481v1 Announce Type: new Abstract: Crowdsourced pairwise evaluation has emerged as a scalable approach for assessing foundation models. However, applying it to Text to Speech(TTS) introdu

researcharxiv-cs-cl
24 Apr 2026
Research

Propensity Inference: Environmental Contributors to LLM Behaviour

DGX agent

arXiv:2604.21098v1 Announce Type: new Abstract: Motivated by loss of control risks from misaligned AI systems, we develop and apply methods for measuring language models' propensity for unsanctioned b

researcharxiv-cs-ai
24 Apr 2026
Model Releases

Reasoning About Traversability: Language-Guided Off-Road 3D Trajectory Planning

DGX agent

arXiv:2604.21249v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) enable high-level semantic reasoning for end-to-end autonomous driving, particularly in unstructured environments, e

model-releasesarxiv-cs-ro
24 Apr 2026
Model Releases

Reinforcing privacy reasoning in LLMs via normative simulacra from fiction

DGX agent

arXiv:2604.20904v1 Announce Type: cross Abstract: Information handling practices of LLM agents are broadly misaligned with the contextual privacy expectations of their users. Contextual Integrity (CI)

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Remember o3 was only a year and a week ago! Also, only GPT-5.5 seemed to take the 'evolution' piece seriously and change the setting rather …

DGX agent

I cannot provide an accurate summary for this entry as the text appears incomplete and lacks sufficient context. The post fragment references o3 (likely an AI model), GPT-5.5, and discusses timeline/e

model-releasesethan-mollick--x
24 Apr 2026
Research

Sink-Token-Aware Pruning for Fine-Grained Video Understanding in Efficient Video LLMs

DGX agent

arXiv:2604.20937v1 Announce Type: new Abstract: Video Large Language Models (Video LLMs) incur high inference latency due to a large number of visual tokens provided to LLMs. To address this, training

researcharxiv-cs-lg
24 Apr 2026
Model Releases

SparseGF: A Height-Aware Sparse Segmentation Framework with Context Compression for Robust Ground Filtering Across Urban to Natural Scenes

DGX agent

arXiv:2604.21356v1 Announce Type: new Abstract: High-quality digital terrain models derived from airborne laser scanning (ALS) data are essential for a wide range of geospatial analyses, and their gen

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Strategic Heterogeneous Multi-Agent Architecture for Cost-Effective Code Vulnerability Detection

DGX agent

arXiv:2604.21282v1 Announce Type: cross Abstract: Automated code vulnerability detection is critical for software security, yet existing approaches face a fundamental trade-off between detection accur

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

The Coding Assistant Breakdown: More Tokens Please

DGX agent

This analysis examines the token consumption and economics of coding assistants, likely comparing different AI models' efficiency and cost-effectiveness for code generation tasks. The piece probably d

model-releasessemianalysis
24 Apr 2026
Model Releases

The Root Theorem of Context Engineering

DGX agent

arXiv:2604.20874v1 Announce Type: cross Abstract: Every system that maintains a large language model conversation beyond a single session faces two inescapable constraints: the context window is finit

model-releasesarxiv-cs-cl
24 Apr 2026
Research

TRACES: Tagging Reasoning Steps for Adaptive Cost-Efficient Early-Stopping

DGX agent

arXiv:2604.21057v1 Announce Type: new Abstract: The field of Language Reasoning Models (LRMs) has been very active over the past few years with advances in training and inference techniques enabling L

researcharxiv-cs-cl
24 Apr 2026
Research

Transferable SCF-Acceleration through Solver-Aligned Initialization Learning

DGX agent

arXiv:2604.21657v1 Announce Type: new Abstract: Machine learning methods that predict initial guesses from molecular geometry can reduce this cost, but matrix-prediction models fail when extrapolating

researcharxiv-cs-lg
24 Apr 2026
Research

VVS: Accelerating Speculative Decoding for Visual Autoregressive Generation via Partial Verification Skipping

DGX agent

arXiv:2511.13587v2 Announce Type: replace-cross Abstract: Visual autoregressive (AR) generation models have demonstrated strong potential for image generation, yet their next-token-prediction paradigm

researcharxiv-cs-ai
24 Apr 2026
Model Releases

We benchmarked GPT-5.5 on document understanding 📄📊 We ran it through ParseBench, our comprehensive OCR benchmark over enterprise document…

DGX agent

We benchmarked GPT-5.5 on document understanding 📄📊 We ran it through ParseBench, our comprehensive OCR benchmark over enterprise documents. We evaluated metrics across various dimensions: visual grou

model-releasesjerry-liu--x
24 Apr 2026
Model Releases

When to Trust the Answer: Question-Aligned Semantic Nearest Neighbor Entropy for Safer Surgical VQA

DGX agent

arXiv:2511.01458v2 Announce Type: replace-cross Abstract: Safety and reliability are critical for deploying visual question answering (VQA) systems in surgery, where incorrect or ambiguous responses c

model-releasesarxiv-cs-ai
24 Apr 2026
← Previous
1…561562563564565…1395
Next →