AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,597 results
21 Apr 2026

This is what I’ve been cooking in the past 4 months . GPT Image 2 is over a massive 240 elo jump over the second place model, marking the bi…

TutorialsDGX agent

This is what I’ve been cooking in the past 4 months . GPT Image 2 is over a massive 240 elo jump over the second place model, marking the biggest jump bigger than the rest of the leaderboard combined

Too Correct to Learn: Reinforcement Learning on Saturated Reasoning Data

Model ReleasesDGX agent

arXiv:2604.18493v1 Announce Type: new Abstract: Reinforcement Learning (RL) enhances LLM reasoning, yet a paradox emerges as models scale: strong base models saturate standard benchmarks (e.g., MATH),

Towards Real-Time ECG and EMG Modeling on mu NPUs

ResearchDGX agent

arXiv:2604.18067v1 Announce Type: new Abstract: The miniaturisation of neural processing units (NPUs) and other low-power accelerators has enabled their integration into microcontroller-scale wearable

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
20 Apr 2026

Arrow 1.1, the latest model from @QuiverAI, is now supported in ComfyUI as a Partner Node, offering scalable, editable SVG output directly w…

ApplicationsDGX agent

Arrow 1.1, the latest model from @QuiverAI, is now supported in ComfyUI as a Partner Node, offering scalable, editable SVG output directly within ComfyUI. What creators use it for: - Logos & wordmarks

Curing Miracle Steps in LLM Mathematical Reasoning with Rubric Rewards

ResearchDGX agent

arXiv:2510.07774v3 Announce Type: replace Abstract: In this paper, we observe that current models are susceptible to reward hacking, leading to a substantial overestimation of a model's reasoning abil

DriveLaW:Unifying Planning and Video Generation in a Latent Driving World

Model ReleasesDGX agent

arXiv:2512.23421v3 Announce Type: replace Abstract: World models have become crucial for autonomous driving, as they learn how scenarios evolve over time to address the long-tail challenges of the rea

Early Detection of Acute Myeloid Leukemia (AML) Using YOLOv12 Deep Learning Model

ResearchDGX agent

arXiv:2604.16082v1 Announce Type: cross Abstract: Acute Myeloid Leukemia (AML) is one of the most life-threatening type of blood cancers, and its accurate classification is considered and remains a ch

Follow the Flow: On Information Flow Across Textual Tokens in Text-to-Image Models

SafetyDGX agent

arXiv:2504.01137v3 Announce Type: replace Abstract: Text-to-image generation models suffer from alignment problems, where generated images fail to accurately capture the objects and relations in the t

Import AI 454: Automating alignment research; safety study of a Chinese model; HiFloat4

SafetyDGX agent

This newsletter covers three main topics: advances in automating alignment research to improve AI safety processes, a safety evaluation study of a Chinese AI model, and technical details about HiFloat

Prototype-Grounded Concept Models for Verifiable Concept Alignment

SafetyDGX agent

arXiv:2604.16076v1 Announce Type: cross Abstract: Concept Bottleneck Models (CBMs) aim to improve interpretability in Deep Learning by structuring predictions through human-understandable concepts, bu

The Spectral Geometry of Thought: Phase Transitions, Instruction Reversal, Token-Level Dynamics, and Perfect Correctness Prediction in How Transformers Reason

Model ReleasesDGX agent

arXiv:2604.15350v1 Announce Type: new Abstract: We discover that large language models exhibit spectral phase transitions in their hidden activation spaces when engaging in reasoning versus factual re

Where Do Vision-Language Models Fail? World Scale Analysis for Image Geolocalization

Local AiDGX agent

arXiv:2604.16248v1 Announce Type: new Abstract: Image geolocalization has traditionally been addressed through retrieval-based place recognition or geometry-based visual localization pipelines. Recent

19 Apr 2026

Sources: Google is in talks with Marvell Technology to develop a memory processing unit that works alongside TPUs, and a new TPU for running AI models (Qianer Liu/The Information)

HardwareDGX agent

Qianer Liu / The Information: Sources: Google is in talks with Marvell Technology to develop a memory processing unit that works alongside TPUs, and a new TPU for running AI models — Google is in talk

17 Apr 2026

A new programming model for durable execution

ToolsDGX agent

This article from Vercel presents a new programming model designed to enable durable execution of applications, likely addressing how to handle long-running tasks, failures, and retries in serverless

Attribution, Citation, and Quotation: A Survey of Evidence-based Text Generation with Large Language Models

ResearchDGX agent

arXiv:2508.15396v2 Announce Type: replace Abstract: The increasing adoption of large language models (LLMs) has raised serious concerns about their reliability and trustworthiness. As a result, a grow

Bit-Accurate Modeling of GPU Matrix Multiply-Accumulate Units: Demystifying Numerical Discrepancy and Accuracy

HardwareDGX agent

arXiv:2511.10909v2 Announce Type: replace-cross Abstract: Modern AI accelerators rely on matrix multiply-accumulate units (MMAUs), such as NVIDIA Tensor Cores and AMD Matrix Cores, to accelerate deep

Blinded Multi-Rater Comparative Evaluation of a Large Language Model and Clinician-Authored Responses in CGM-Informed Diabetes Counseling

SafetyDGX agent

arXiv:2604.15124v1 Announce Type: new Abstract: Continuous glucose monitoring (CGM) is central to diabetes care, but explaining CGM patterns clearly and empathetically remains time-intensive. Evidence

Can Large Language Models Detect Methodological Flaws? Evidence from Gesture Recognition for UAV-Based Rescue Operation Based on Deep Learning

ApplicationsDGX agent

arXiv:2604.14161v1 Announce Type: new Abstract: Reliable evaluation is essential in machine learning research, yet methodological flaws-particularly data leakage-continue to undermine the validity of

Frame forecasting in cine MRI using the PCA respiratory motion model: comparing recurrent neural networks trained online and transformers

Model ReleasesDGX agent

arXiv:2410.05882v3 Announce Type: replace-cross Abstract: Respiratory motion complicates accurate irradiation of thoraco-abdominal tumors during radiotherapy, as treatment-system latency entails targe

Language Model Fine-Tuning on Scaled Survey Data for Predicting Distributions of Public Opinions

ApplicationsDGX agent

arXiv:2502.16761v2 Announce Type: replace Abstract: Large language models (LLMs) present novel opportunities in public opinion research by predicting survey responses in advance during the early stage

MapSR: Prompt-Driven Land Cover Map Super-Resolution via Vision Foundation Models

ResearchDGX agent

arXiv:2604.14582v1 Announce Type: new Abstract: High-resolution (HR) land-cover mapping is often constrained by the high cost of dense HR annotations. We revisit this problem from the perspective of m

MemGround: Long-Term Memory Evaluation Kit for Large Language Models in Gamified Scenarios

Model ReleasesDGX agent

arXiv:2604.14158v1 Announce Type: new Abstract: Current evaluations of long-term memory in LLMs are fundamentally static. By fixating on simple retrieval and short-context inference, they neglect the

Mitigating LLM biases toward spurious social contexts using direct preference optimization

Model ReleasesDGX agent

arXiv:2604.02585v2 Announce Type: replace-cross Abstract: LLMs are increasingly used for high-stakes decision-making, yet their sensitivity to spurious contextual information can introduce harmful bia

Model-Based Reinforcement Learning under Random Observation Delays

ApplicationsDGX agent

arXiv:2509.20869v2 Announce Type: replace Abstract: Delays frequently occur in real-world environments, yet standard reinforcement learning (RL) algorithms often assume instantaneous perception of the

Model Capability Dominates: Inference-Time Optimization Lessons from AIMO 3

HardwareDGX agent

arXiv:2603.27844v2 Announce Type: replace Abstract: Majority voting over multiple LLM attempts improves mathematical reasoning, but correlated errors limit the effective sample size. A natural fix is

SegviGen: Repurposing 3D Generative Model for Part Segmentation

ResearchDGX agent

arXiv:2603.16869v2 Announce Type: replace Abstract: We introduce SegviGen, a framework that repurposes native 3D generative models for 3D part segmentation. Existing pipelines either lift strong 2D pr

16 Apr 2026

Atelier: a canvas for thinking and making with local models.

Local AiDGX agent

Atelier is a canvas-like system that leverages generative image and video models to blend spaces for thinking and creation, where both references and generated assets co-exist in one unified workspace

BenGER: A Collaborative Web Platform for End-to-End Benchmarking of German Legal Tasks

Model ReleasesDGX agent

arXiv:2604.13583v1 Announce Type: new Abstract: Evaluating large language models (LLMs) for legal reasoning requires workflows that span task design, expert annotation, model execution, and metric-bas

Hybrid Attention Model Using Feature Decomposition and Knowledge Distillation for Glucose Forecasting

ApplicationsDGX agent

arXiv:2411.10703v3 Announce Type: replace Abstract: The availability of continuous glucose monitors as over-the-counter commodities have created a unique opportunity to monitor a person's blood glucos

Memo: the White House notified Cabinet departments that OMB is setting up protections that would allow their agencies to begin using Anthropic's Mythos model (Bloomberg)

IndustryDGX agent

Bloomberg: Memo: the White House notified Cabinet departments that OMB is setting up protections that would allow their agencies to begin using Anthropic's Mythos model — The US government is preparin

Numerical Instability and Chaos: Quantifying the Unpredictability of Large Language Models

AgentsDGX agent

arXiv:2604.13206v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are increasingly integrated into agentic workflows, their unpredictability stemming from numerical instability has eme

PRiMeFlow: Capturing Complex Expression Heterogeneity in Perturbation Response Modelling

ResearchDGX agent

arXiv:2604.13986v1 Announce Type: new Abstract: Predicting the effects of perturbations in-silico on cell state can identify drivers of cell behavior at scale and accelerate drug discovery. However, m

Qwen3.6-35B-A3B on my laptop drew me a better pelican than Claude Opus 4.7

Model ReleasesDGX agent

For anyone who has been (inadvisably) taking my pelican riding a bicycle benchmark seriously as a robust way to test models, here are pelicans from this morning's two big model releases - Qwen3.6-35B-

ROSE: Retrieval-Oriented Segmentation Enhancement

Model ReleasesDGX agent

arXiv:2604.14147v1 Announce Type: new Abstract: Existing segmentation models based on multimodal large language models (MLLMs), such as LISA, often struggle with novel or emerging entities due to thei

Stochastic Trust-Region Methods for Over-parameterized Models

Model ReleasesDGX agent

arXiv:2604.14017v1 Announce Type: cross Abstract: Under interpolation-type assumptions such as the strong growth condition, stochastic optimization methods can attain convergence rates comparable to f

Synthesis: Arxiv-Cs-Ai

SynthesesDGX agent

Auto-generated synthesis of 1623 entries about arxiv-cs-ai

Training-Free Test-Time Contrastive Learning for Large Language Models

AgentsDGX agent

arXiv:2604.13552v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate strong reasoning capabilities, but their performance often degrades under distribution shift. Existing test-tim

UHR-BAT: Budget-Aware Token Compression Vision-Language model for Ultra-High-Resolution Remote Sensing

ResearchDGX agent

arXiv:2604.13565v1 Announce Type: new Abstract: Ultra-high-resolution (UHR) remote sensing imagery couples kilometer-scale context with query-critical evidence that may occupy only a few pixels. Such

15 Apr 2026

Cool stuff Google Cloud customers built, April edition: BMW big on SLMs, MLB’s Scout Insights AI, personalized resort experiences

Model ReleasesDGX agent

AI and cloud technology are reshaping every corner of every industry around the world. Without our customers, who are building the future on our platform, there would be no Google Cloud. In this regul

Detecting Complex Money Laundering Patterns with Incremental and Distributed Graph Modeling

Model ReleasesDGX agent

arXiv:2604.01315v2 Announce Type: replace Abstract: Money launderers take advantage of limitations in existing detection approaches by hiding their financial footprints in a deceitful manner. They man

Enabling Ultra-Fast Cardiovascular Imaging Across Heterogeneous Clinical Environments with A Generalist Foundation Model and Multimodal Database

ResearchDGX agent

arXiv:2512.21652v2 Announce Type: replace-cross Abstract: Multimodal cardiovascular magnetic resonance (CMR) imaging provides comprehensive and non-invasive insights into cardiovascular disease (CVD)

Gemini 3.1 Flash TTS: the next generation of expressive AI speech

Model ReleasesDGX agent

Google DeepMind's Gemini 3.1 Flash TTS is a text-to-speech model delivering improved controllability, expressivity, and quality for developers, enterprises, and everyday users building AI-speech appli

Generative Modeling Enables Molecular Structure Retrieval from Coulomb Explosion Imaging

ResearchDGX agent

arXiv:2511.00179v2 Announce Type: replace-cross Abstract: Capturing the structural changes that molecules undergo during chemical reactions in real space and time is a long-standing dream and an essen

Gradient boundaries through confidence intervals for forced alignment estimates using model ensembles

SafetyDGX agent

arXiv:2506.01256v4 Announce Type: replace-cross Abstract: Forced alignment is a common tool to align audio with orthographic and phonetic transcriptions. Most forced alignment tools provide only point

IAD-Unify: A Region-Grounded Unified Model for Industrial Anomaly Segmentation, Understanding, and Generation

Model ReleasesDGX agent

arXiv:2604.12440v1 Announce Type: cross Abstract: Real-world industrial inspection requires not only localizing defects, but also explaining them in natural language and generating controlled defect e

Is it possible for an open-source AI that you run at home to become as powerful as that of chatgpt and others at that level?

IndustryDGX agent

This Reddit thread from r/ChatGPT discusses whether locally-run open-source AI models can match the capabilities of frontier models like ChatGPT. Leading open-source models like Llama 3.3 70B and Deep

LASA: Language-Agnostic Semantic Alignment at the Semantic Bottleneck for LLM Safety

Model ReleasesDGX agent

arXiv:2604.12710v1 Announce Type: cross Abstract: Large language models (LLMs) often demonstrate strong safety performance in high-resource languages, yet exhibit severe vulnerabilities when queried i

Olmo 3

Model ReleasesDGX agent

arXiv:2512.13961v2 Announce Type: replace Abstract: We introduce Olmo 3, a family of state-of-the-art, fully-open language models at the 7B and 32B parameter scales. Olmo 3 model construction targets

Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe

SafetyDGX agent

arXiv:2604.13016v1 Announce Type: cross Abstract: On-policy distillation (OPD) has become a core technique in the post-training of large language models, yet its training dynamics remain poorly unders

StableSketcher: Enhancing Diffusion Model for Pixel-based Sketch Generation via Visual Question Answering Feedback

SafetyDGX agent

arXiv:2510.20093v2 Announce Type: replace-cross Abstract: Although recent advancements in diffusion models have significantly enriched the quality of generated images, challenges remain in synthesizin

Uncertainty Quantification on Graph Learning: A Survey

ApplicationsDGX agent

arXiv:2404.14642v4 Announce Type: replace Abstract: Graphical models have demonstrated their exceptional capabilities across numerous applications. However, their performance, confidence, and trustwor

14 Apr 2026

3D-Anchored Lookahead Planning for Persistent Robotic Scene Memory via World-Model-Based MCTS

ResearchDGX agent

arXiv:2604.11302v1 Announce Type: cross Abstract: We present 3D-Anchored Lookahead Planning (3D-ALP), a System 2 reasoning engine for robotic manipulation that combines Monte Carlo Tree Search (MCTS)

A graph of headcount trends in high AI-exposure jobs for early in career (age 22-25) annotated with big AI model releases. Among workers age…

IndustryDGX agent

A graph of headcount trends in high AI-exposure jobs for early in career (age 22-25) annotated with big AI model releases. Among workers age 22-25, employment in the most AI-exposed occupations has fa

A Progressive Training Strategy for Vision-Language Models to Counteract Spatio-Temporal Hallucinations in Embodied Reasoning

ResearchDGX agent

arXiv:2604.10506v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have made significant strides in static image understanding but continue to face critical hurdles in spatiotemporal reason

Aligning What LLMs Do and Say: Towards Self-Consistent Explanations

Model ReleasesDGX agent

arXiv:2506.07523v3 Announce Type: replace Abstract: Large language models (LLMs) seem to offer an easy path to interpretability: just ask them to explain their answers. Yet the features driving an ans

Attention Sinks as Internal Signals for Hallucination Detection in Large Language Models

ResearchDGX agent

arXiv:2604.10697v1 Announce Type: new Abstract: Large language models frequently exhibit hallucinations: fluent and confident outputs that are factually incorrect or unsupported by the input context.

Beyond Model Design: Data-Centric Training and Self-Ensemble for Gaussian Color Image Denoising

Local AiDGX agent

arXiv:2604.11468v1 Announce Type: new Abstract: This paper presents our solution to the NTIRE 2026 Image Denoising Challenge (Gaussian color image denoising at fixed noise level sigma = 50). Rather th

CAGenMol: Condition-Aware Diffusion Language Model for Goal-Directed Molecular Generation

SafetyDGX agent

arXiv:2604.11483v1 Announce Type: new Abstract: Goal-directed molecular generation requires satisfying heterogeneous constraints such as protein--ligand compatibility and multi-objective drug-like pro

Cognitive Pivot Points and Visual Anchoring: Unveiling and Rectifying Hallucinations in Multimodal Reasoning Models

Local AiDGX agent

arXiv:2604.10219v1 Announce Type: new Abstract: Multimodal Large Reasoning Models (MLRMs) have achieved remarkable strides in visual reasoning through test time compute scaling, yet long chain reasoni

COSMIK-MPPI: Scaling Constrained Model Predictive Control to Collision Avoidance in Close-Proximity Dynamic Human Environments

SafetyDGX agent

arXiv:2604.10358v1 Announce Type: new Abstract: Ensuring safe physical interaction between torque-controlled manipulators and humans is essential for deploying robots in everyday environments. Model P

← Previous
1…192193194195196…1010
Next →