AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlog
85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,047 results
Model Releases

Language Models Don't Know What You Want: Evaluating Personalization in Deep Research Needs Real Users

DGX agent

arXiv:2603.16120v2 Announce Type: replace Abstract: Deep Research (DR) systems help researchers cope with ballooning publishing counts. Such tools synthesize scientific papers to answer research queri

model-releasesarxiv-cs-cl
21 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Latent Preference Modeling for Cross-Session Personalized Tool Calling

DGX agent

arXiv:2604.17886v1 Announce Type: new Abstract: Users often omit essential details in their requests to LLM-based agents, resulting in under-specified inputs for tool use. This poses a fundamental cha

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Long-Text-to-Image Generation via Compositional Prompt Decomposition

DGX agent

arXiv:2604.18258v1 Announce Type: new Abstract: While modern text-to-image (T2I) models excel at generating images from intricate prompts, they struggle to capture the key details when the inputs are

model-releasesarxiv-cs-cv
21 Apr 2026
Research

LoReC: Rethinking Large Language Models for Graph Data Analysis

DGX agent

arXiv:2604.17897v1 Announce Type: new Abstract: The advent of Large Language Models (LLMs) has fundamentally reshaped the way we interact with graphs, giving rise to a new paradigm called GraphLLM. As

researcharxiv-cs-lg
21 Apr 2026
Research

Modeling, Control and Self-sensing of Dielectric Elastomer Soft Actuators: A Review

DGX agent

arXiv:2604.17199v1 Announce Type: new Abstract: Dielectric elastomer actuators (DEAs) have garnered extensive attention especially in soft robotic applications over the past few decades owing to the a

researcharxiv-cs-ro
21 Apr 2026
Model Releases

Modeling Higher-Order Brain Interactions via a Multi-View Information Bottleneck Framework for fMRI-based Psychiatric Diagnosis

DGX agent

arXiv:2604.17713v1 Announce Type: new Abstract: Resting-state functional magnetic resonance imaging (fMRI) has emerged as a cornerstone for psychiatric diagnosis, yet most approaches rely on pairwise

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Modelling Gas-Phase Reaction Kinetics with Guided Particle Diffusion Sampling

DGX agent

arXiv:2604.16461v1 Announce Type: cross Abstract: Physics-guided sampling with diffusion priors has recently shown strong performance in solving complex systems of partial differential equations (PDEs

model-releasesarxiv-cs-lg
21 Apr 2026
Safety

Operationalizing Fairness in Text-to-Image Models: A Survey of Bias, Fairness Audits and Mitigation Strategies

DGX agent

arXiv:2604.16516v1 Announce Type: new Abstract: Text-to-Image (T2I) generation models have been widely adopted across various industries, yet are criticized for frequently exhibiting societal stereoty

safetyarxiv-cs-cv
21 Apr 2026
Model Releases

PiCa: Parameter-Efficient Fine-Tuning with Column Space Projection

DGX agent

arXiv:2505.20211v3 Announce Type: replace Abstract: Fine-tuning large foundation models is essential for building expert models tailored to specialized tasks and domains, but fully updating billions o

model-releasesarxiv-cs-lg
21 Apr 2026
Safety

Safer Trajectory Planning with CBF-guided Diffusion Model for Unmanned Aerial Vehicles

DGX agent

arXiv:2604.17527v1 Announce Type: new Abstract: Safe and agile trajectory planning is essential for autonomous systems, especially during complex aerobatic maneuvers. Motivated by the recent success o

safetyarxiv-cs-ro
21 Apr 2026
Applications

Shifting the Gradient: Understanding How Defensive Training Methods Protect Language Model Integrity

DGX agent

arXiv:2604.16423v1 Announce Type: new Abstract: Defensive training methods such as positive preventative steering (PPS) and inoculation prompting (IP) offer surprising results through seemingly simila

applicationsarxiv-cs-lg
21 Apr 2026
Model Releases

SinkRouter: Sink-Aware Routing for Efficient Long-Context Decoding in Large Language and Multimodal Models

DGX agent

arXiv:2604.16883v1 Announce Type: new Abstract: In long-context decoding for LLMs and LMMs, attention becomes increasingly memory-bound because each decoding step must load a large amount of KV-cache

model-releasesarxiv-cs-lg
21 Apr 2026
Research

Spotlights and Blindspots: Evaluation Machine-Generated Text Detection

DGX agent

arXiv:2604.16607v1 Announce Type: new Abstract: With the rise of generative language models, machine-generated text detection has become a critical challenge. A wide variety of models is available, bu

researcharxiv-cs-cl
21 Apr 2026
Research

Target Parameterization in Diffusion Models for Nonlinear Spatiotemporal System Identification

DGX agent

arXiv:2604.17566v1 Announce Type: cross Abstract: Machine learning is becoming increasingly important for nonlinear system identification, including dynamical systems with spatially distributed output

researcharxiv-cs-lg
21 Apr 2026
Tutorials

This is what I’ve been cooking in the past 4 months . GPT Image 2 is over a massive 240 elo jump over the second place model, marking the bi…

DGX agent

This is what I’ve been cooking in the past 4 months . GPT Image 2 is over a massive 240 elo jump over the second place model, marking the biggest jump bigger than the rest of the leaderboard combined

tutorialsjeremy-howard--x
21 Apr 2026
Model Releases

Too Correct to Learn: Reinforcement Learning on Saturated Reasoning Data

DGX agent

arXiv:2604.18493v1 Announce Type: new Abstract: Reinforcement Learning (RL) enhances LLM reasoning, yet a paradox emerges as models scale: strong base models saturate standard benchmarks (e.g., MATH),

model-releasesarxiv-cs-lg
21 Apr 2026
Research

Towards Real-Time ECG and EMG Modeling on mu NPUs

DGX agent

arXiv:2604.18067v1 Announce Type: new Abstract: The miniaturisation of neural processing units (NPUs) and other low-power accelerators has enabled their integration into microcontroller-scale wearable

researcharxiv-cs-lg
21 Apr 2026
Applications

Arrow 1.1, the latest model from @QuiverAI, is now supported in ComfyUI as a Partner Node, offering scalable, editable SVG output directly w…

DGX agent

Arrow 1.1, the latest model from @QuiverAI, is now supported in ComfyUI as a Partner Node, offering scalable, editable SVG output directly within ComfyUI. What creators use it for: - Logos & wordmarks

applicationscomfyui--x
20 Apr 2026
Research

Curing Miracle Steps in LLM Mathematical Reasoning with Rubric Rewards

DGX agent

arXiv:2510.07774v3 Announce Type: replace Abstract: In this paper, we observe that current models are susceptible to reward hacking, leading to a substantial overestimation of a model's reasoning abil

researcharxiv-cs-cl
20 Apr 2026
Model Releases

DriveLaW:Unifying Planning and Video Generation in a Latent Driving World

DGX agent

arXiv:2512.23421v3 Announce Type: replace Abstract: World models have become crucial for autonomous driving, as they learn how scenarios evolve over time to address the long-tail challenges of the rea

model-releasesarxiv-cs-cv
20 Apr 2026
Research

Early Detection of Acute Myeloid Leukemia (AML) Using YOLOv12 Deep Learning Model

DGX agent

arXiv:2604.16082v1 Announce Type: cross Abstract: Acute Myeloid Leukemia (AML) is one of the most life-threatening type of blood cancers, and its accurate classification is considered and remains a ch

researcharxiv-cs-ai
20 Apr 2026
Safety

Follow the Flow: On Information Flow Across Textual Tokens in Text-to-Image Models

DGX agent

arXiv:2504.01137v3 Announce Type: replace Abstract: Text-to-image generation models suffer from alignment problems, where generated images fail to accurately capture the objects and relations in the t

safetyarxiv-cs-cl
20 Apr 2026
Safety

Import AI 454: Automating alignment research; safety study of a Chinese model; HiFloat4

DGX agent

This newsletter covers three main topics: advances in automating alignment research to improve AI safety processes, a safety evaluation study of a Chinese AI model, and technical details about HiFloat

safetyimport-ai
20 Apr 2026
Safety

Prototype-Grounded Concept Models for Verifiable Concept Alignment

DGX agent

arXiv:2604.16076v1 Announce Type: cross Abstract: Concept Bottleneck Models (CBMs) aim to improve interpretability in Deep Learning by structuring predictions through human-understandable concepts, bu

safetyarxiv-cs-ai
20 Apr 2026
Model Releases

The Spectral Geometry of Thought: Phase Transitions, Instruction Reversal, Token-Level Dynamics, and Perfect Correctness Prediction in How Transformers Reason

DGX agent

arXiv:2604.15350v1 Announce Type: new Abstract: We discover that large language models exhibit spectral phase transitions in their hidden activation spaces when engaging in reasoning versus factual re

model-releasesarxiv-cs-lg
20 Apr 2026
Local Ai

Where Do Vision-Language Models Fail? World Scale Analysis for Image Geolocalization

DGX agent

arXiv:2604.16248v1 Announce Type: new Abstract: Image geolocalization has traditionally been addressed through retrieval-based place recognition or geometry-based visual localization pipelines. Recent

local-aiarxiv-cs-cv
20 Apr 2026
Hardware

Sources: Google is in talks with Marvell Technology to develop a memory processing unit that works alongside TPUs, and a new TPU for running AI models (Qianer Liu/The Information)

DGX agent

Qianer Liu / The Information: Sources: Google is in talks with Marvell Technology to develop a memory processing unit that works alongside TPUs, and a new TPU for running AI models — Google is in talk

hardwaretechmeme
19 Apr 2026
Tools

A new programming model for durable execution

DGX agent

This article from Vercel presents a new programming model designed to enable durable execution of applications, likely addressing how to handle long-running tasks, failures, and retries in serverless

toolsvercel-blog
17 Apr 2026
Research

Attribution, Citation, and Quotation: A Survey of Evidence-based Text Generation with Large Language Models

DGX agent

arXiv:2508.15396v2 Announce Type: replace Abstract: The increasing adoption of large language models (LLMs) has raised serious concerns about their reliability and trustworthiness. As a result, a grow

researcharxiv-cs-cl
17 Apr 2026
Hardware

Bit-Accurate Modeling of GPU Matrix Multiply-Accumulate Units: Demystifying Numerical Discrepancy and Accuracy

DGX agent

arXiv:2511.10909v2 Announce Type: replace-cross Abstract: Modern AI accelerators rely on matrix multiply-accumulate units (MMAUs), such as NVIDIA Tensor Cores and AMD Matrix Cores, to accelerate deep

hardwarearxiv-cs-lg
17 Apr 2026
Safety

Blinded Multi-Rater Comparative Evaluation of a Large Language Model and Clinician-Authored Responses in CGM-Informed Diabetes Counseling

DGX agent

arXiv:2604.15124v1 Announce Type: new Abstract: Continuous glucose monitoring (CGM) is central to diabetes care, but explaining CGM patterns clearly and empathetically remains time-intensive. Evidence

safetyarxiv-cs-cl
17 Apr 2026
Applications

Can Large Language Models Detect Methodological Flaws? Evidence from Gesture Recognition for UAV-Based Rescue Operation Based on Deep Learning

DGX agent

arXiv:2604.14161v1 Announce Type: new Abstract: Reliable evaluation is essential in machine learning research, yet methodological flaws-particularly data leakage-continue to undermine the validity of

applicationsarxiv-cs-cl
17 Apr 2026
Model Releases

Frame forecasting in cine MRI using the PCA respiratory motion model: comparing recurrent neural networks trained online and transformers

DGX agent

arXiv:2410.05882v3 Announce Type: replace-cross Abstract: Respiratory motion complicates accurate irradiation of thoraco-abdominal tumors during radiotherapy, as treatment-system latency entails targe

model-releasesarxiv-cs-cv
17 Apr 2026
Applications

Language Model Fine-Tuning on Scaled Survey Data for Predicting Distributions of Public Opinions

DGX agent

arXiv:2502.16761v2 Announce Type: replace Abstract: Large language models (LLMs) present novel opportunities in public opinion research by predicting survey responses in advance during the early stage

applicationsarxiv-cs-cl
17 Apr 2026
Research

MapSR: Prompt-Driven Land Cover Map Super-Resolution via Vision Foundation Models

DGX agent

arXiv:2604.14582v1 Announce Type: new Abstract: High-resolution (HR) land-cover mapping is often constrained by the high cost of dense HR annotations. We revisit this problem from the perspective of m

researcharxiv-cs-cv
17 Apr 2026
Model Releases

MemGround: Long-Term Memory Evaluation Kit for Large Language Models in Gamified Scenarios

DGX agent

arXiv:2604.14158v1 Announce Type: new Abstract: Current evaluations of long-term memory in LLMs are fundamentally static. By fixating on simple retrieval and short-context inference, they neglect the

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Mitigating LLM biases toward spurious social contexts using direct preference optimization

DGX agent

arXiv:2604.02585v2 Announce Type: replace-cross Abstract: LLMs are increasingly used for high-stakes decision-making, yet their sensitivity to spurious contextual information can introduce harmful bia

model-releasesarxiv-cs-cl
17 Apr 2026
Applications

Model-Based Reinforcement Learning under Random Observation Delays

DGX agent

arXiv:2509.20869v2 Announce Type: replace Abstract: Delays frequently occur in real-world environments, yet standard reinforcement learning (RL) algorithms often assume instantaneous perception of the

applicationsarxiv-cs-lg
17 Apr 2026
Hardware

Model Capability Dominates: Inference-Time Optimization Lessons from AIMO 3

DGX agent

arXiv:2603.27844v2 Announce Type: replace Abstract: Majority voting over multiple LLM attempts improves mathematical reasoning, but correlated errors limit the effective sample size. A natural fix is

hardwarearxiv-cs-cl
17 Apr 2026
Research

SegviGen: Repurposing 3D Generative Model for Part Segmentation

DGX agent

arXiv:2603.16869v2 Announce Type: replace Abstract: We introduce SegviGen, a framework that repurposes native 3D generative models for 3D part segmentation. Existing pipelines either lift strong 2D pr

researcharxiv-cs-cv
17 Apr 2026
Local Ai

Atelier: a canvas for thinking and making with local models.

DGX agent

Atelier is a canvas-like system that leverages generative image and video models to blend spaces for thinking and creation, where both references and generated assets co-exist in one unified workspace

local-air-stablediffusion
16 Apr 2026
Model Releases

BenGER: A Collaborative Web Platform for End-to-End Benchmarking of German Legal Tasks

DGX agent

arXiv:2604.13583v1 Announce Type: new Abstract: Evaluating large language models (LLMs) for legal reasoning requires workflows that span task design, expert annotation, model execution, and metric-bas

model-releasesarxiv-cs-cl
16 Apr 2026
Applications

Hybrid Attention Model Using Feature Decomposition and Knowledge Distillation for Glucose Forecasting

DGX agent

arXiv:2411.10703v3 Announce Type: replace Abstract: The availability of continuous glucose monitors as over-the-counter commodities have created a unique opportunity to monitor a person's blood glucos

applicationsarxiv-cs-lg
16 Apr 2026
Industry

Memo: the White House notified Cabinet departments that OMB is setting up protections that would allow their agencies to begin using Anthropic's Mythos model (Bloomberg)

DGX agent

Bloomberg: Memo: the White House notified Cabinet departments that OMB is setting up protections that would allow their agencies to begin using Anthropic's Mythos model — The US government is preparin

industrytechmeme
16 Apr 2026
Agents

Numerical Instability and Chaos: Quantifying the Unpredictability of Large Language Models

DGX agent

arXiv:2604.13206v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are increasingly integrated into agentic workflows, their unpredictability stemming from numerical instability has eme

agentsarxiv-cs-lg
16 Apr 2026
Research

PRiMeFlow: Capturing Complex Expression Heterogeneity in Perturbation Response Modelling

DGX agent

arXiv:2604.13986v1 Announce Type: new Abstract: Predicting the effects of perturbations in-silico on cell state can identify drivers of cell behavior at scale and accelerate drug discovery. However, m

researcharxiv-cs-lg
16 Apr 2026
Model Releases

Qwen3.6-35B-A3B on my laptop drew me a better pelican than Claude Opus 4.7

DGX agent

For anyone who has been (inadvisably) taking my pelican riding a bicycle benchmark seriously as a robust way to test models, here are pelicans from this morning's two big model releases - Qwen3.6-35B-

model-releasessimon-willison
16 Apr 2026
Model Releases

ROSE: Retrieval-Oriented Segmentation Enhancement

DGX agent

arXiv:2604.14147v1 Announce Type: new Abstract: Existing segmentation models based on multimodal large language models (MLLMs), such as LISA, often struggle with novel or emerging entities due to thei

model-releasesarxiv-cs-cv
16 Apr 2026
← Previous
1…242243244245246…1272
Next →