AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
AllBlog
91,020Total entries
1Added by human
91,019Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,767 results
Safety

Face Density as a Proxy for Data Complexity: Quantifying the Hardness of Instance Count

DGX agent

arXiv:2604.09689v1 Announce Type: cross Abstract: Machine learning progress has historically prioritized model-centric innovations, yet achievable performance is frequently capped by the intrinsic com

safetyarxiv-cs-ai
14 Apr 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

FACT-E: Causality-Inspired Evaluation for Trustworthy Chain-of-Thought Reasoning

DGX agent

arXiv:2604.10693v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting has improved LLM reasoning, but models often generate explanations that appear coherent while containing unfaithful int

safetyarxiv-cs-ai
14 Apr 2026
Safety

FAITH: Factuality Alignment through Integrating Trustworthiness and Honestness

DGX agent

arXiv:2604.10189v1 Announce Type: new Abstract: Large Language Models (LLMs) can generate factually inaccurate content even if they have corresponding knowledge, which critically undermines their reli

safetyarxiv-cs-cl
14 Apr 2026
Local Ai

Forget about VAEs? SenseNova's NEO-unify achieves 31.5 PSNR without an encoder – Native Image Gen is coming.

DGX agent

SenseNova's NEO-unify is an encoder-free unified multimodal model built on a Mixture-of-Transformers (MoT) backbone that eliminates the need for traditional VAE encoders in image generation. The 2B NE

local-air-stablediffusion
14 Apr 2026
Research

From Perception to Planning: Evolving Ego-Centric Task-Oriented Spatiotemporal Reasoning via Curriculum Learning

DGX agent

arXiv:2604.10517v1 Announce Type: new Abstract: Modern vision-language models achieve strong performance in static perception, but remain limited in the complex spatiotemporal reasoning required for e

researcharxiv-cs-ai
14 Apr 2026
Research

GanitLLM: Difficulty-Aware Bengali Mathematical Reasoning through Curriculum-GRPO

DGX agent

arXiv:2601.06767v2 Announce Type: replace-cross Abstract: We present a Bengali mathematical reasoning model called GanitLLM (named after the Bangla word for mathematics, 'Ganit'), together with a new

researcharxiv-cs-ai
14 Apr 2026
Model Releases

GazeVaLM: A Multi-Observer Eye-Tracking Benchmark for Evaluating Clinical Realism in AI-Generated X-Rays

DGX agent

arXiv:2604.11653v1 Announce Type: new Abstract: We introduce GazeVaLM, a public eye-tracking dataset for studying clinical perception during chest radiograph authenticity assessment. The dataset compr

model-releasesarxiv-cs-cv
14 Apr 2026
Safety

GenProve: Learning to Generate Text with Fine-Grained Provenance

DGX agent

arXiv:2601.04932v2 Announce Type: replace Abstract: Large language models (LLM) often hallucinate, and while adding citations is a common solution, it is frequently insufficient for accountability as

safetyarxiv-cs-cl
14 Apr 2026
Model Releases

Governed Reasoning for Institutional AI

DGX agent

arXiv:2604.10658v1 Announce Type: new Abstract: Institutional decisions -- regulatory compliance, clinical triage, prior authorization appeal -- require a different AI architecture than general-purpos

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

GPT vs Claude in a bomberman-style 1v1 game

DGX agent

A Reddit post on r/ChatGPT in which a user built or showcased a Bomberman-style 1v1 game pitting GPT (OpenAI) against Claude (Anthropic) as autonomous AI players, likely using their respective APIs to

model-releasesr-chatgpt
14 Apr 2026
Model Releases

HeceTokenizer: A Syllable-Based Tokenization Approach for Turkish Retrieval

DGX agent

arXiv:2604.10665v1 Announce Type: new Abstract: HeceTokenizer is a syllable-based tokenizer for Turkish that exploits the deterministic six-pattern phonological structure of the language to construct

model-releasesarxiv-cs-cl
14 Apr 2026
Research

How LLMs Might Think

DGX agent

arXiv:2604.09674v1 Announce Type: new Abstract: Do large language models (LLMs) think? Daniel Stoljar and Zhihe Vincent Zhang have recently developed an argument from rationality for the claim that LL

researcharxiv-cs-ai
14 Apr 2026
Model Releases

How You Ask Matters! Adaptive RAG Robustness to Query Variations

DGX agent

arXiv:2604.10745v1 Announce Type: new Abstract: Adaptive Retrieval-Augmented Generation (RAG) promises accuracy and efficiency by dynamically triggering retrieval only when needed and is widely used i

model-releasesarxiv-cs-cl
14 Apr 2026
Research

https://x.com/nousresearch/status/2043969403247616478?s=46

DGX agent

Nous Research shared an announcement or update on their X (formerly Twitter) account, likely relating to their ongoing work in AI model development, fine-tuning, or research releases. Nous Research is

researchnous-research--x
14 Apr 2026
Local Ai

I built a tool that uses diffusion to create user interfaces.

DGX agent

A community developer shared a custom-built tool on r/StableDiffusion that leverages diffusion model technology — typically used for AI image generation — to generate user interface (UI) designs or co

local-air-stablediffusion
14 Apr 2026
Model Releases

I Can't Believe TTA Is Not Better: When Test-Time Augmentation Hurts Medical Image Classification

DGX agent

arXiv:2604.09697v1 Announce Type: cross Abstract: Test-time augmentation (TTA)--aggregating predictions over multiple augmented copies of a test input--is widely assumed to improve classification accu

model-releasesarxiv-cs-ai
14 Apr 2026
Research

I Walk the Line: Examining the Role of Gestalt Continuity in Object Binding for Vision Transformers

DGX agent

arXiv:2604.09942v1 Announce Type: cross Abstract: Object binding is a foundational process in visual cognition, during which low-level perceptual features are joined into object representations. Bindi

researcharxiv-cs-ai
14 Apr 2026
Model Releases

IMPACT: A Dataset for Multi-Granularity Human Procedural Action Understanding in Industrial Assembly

DGX agent

arXiv:2604.10409v1 Announce Type: cross Abstract: We introduce IMPACT, a synchronized five-view RGB-D dataset for deployment-oriented industrial procedural understanding, built around real assembly an

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Improving understanding and trust in AI: How users benefit from interval-based counterfactual explanations

DGX agent

arXiv:2604.09573v1 Announce Type: cross Abstract: Experimental user studies evaluating the effectiveness of different subtypes of post-hoc explanations for black-box models are largely nonexistent. Th

researcharxiv-cs-lg
14 Apr 2026
Local Ai

Inertial Magnetic SLAM Systems Using Low-Cost Sensors

DGX agent

arXiv:2512.10128v2 Announce Type: replace Abstract: Spatially inhomogeneous magnetic fields offer a valuable, non-visual information source for positioning. Among systems leveraging this, magnetic fie

local-aiarxiv-cs-ro
14 Apr 2026
Model Releases

Investigating Bias and Fairness in Appearance-based Gaze Estimation

DGX agent

arXiv:2604.10707v1 Announce Type: new Abstract: While appearance-based gaze estimation has achieved significant improvements in accuracy and domain adaptation, the fairness of these systems across dif

model-releasesarxiv-cs-cv
14 Apr 2026
Tutorials

KL Divergence Between Gaussians: A Step-by-Step Derivation for the Variational Autoencoder Objective

DGX agent

arXiv:2604.11744v1 Announce Type: new Abstract: Kullback-Leibler (KL) divergence is a fundamental concept in information theory that quantifies the discrepancy between two probability distributions. I

tutorialsarxiv-cs-lg
14 Apr 2026
Research

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation

DGX agent

arXiv:2509.20128v2 Announce Type: replace-cross Abstract: Audio-driven facial animation has made significant progress in multimedia applications, with diffusion models showing strong potential for tal

researcharxiv-cs-ai
14 Apr 2026
Applications

Learnable Motion-Focused Tokenization for Effective and Efficient Video Unsupervised Domain Adaptation

DGX agent

arXiv:2604.09955v1 Announce Type: new Abstract: Video Unsupervised Domain Adaptation (VUDA) poses a significant challenge in action recognition, requiring the adaptation of a model from a labeled sour

applicationsarxiv-cs-cv
14 Apr 2026
Agents

Learning to Focus and Precise Cropping: A Reinforcement Learning Framework with Information Gaps and Grounding Loss for MLLMs

DGX agent

arXiv:2603.27494v2 Announce Type: replace-cross Abstract: To enhance the perception and reasoning capabilities of multimodal large language models in complex visual scenes, recent research has introdu

agentsarxiv-cs-ai
14 Apr 2026
Safety

LLM-as-Judge on a Budget

DGX agent

arXiv:2602.15481v2 Announce Type: replace Abstract: LLM-as-a-judge has emerged as a cornerstone technique for evaluating large language models by leveraging LLM reasoning to score prompt-response pair

safetyarxiv-cs-lg
14 Apr 2026
Safety

LLM Nepotism in Organizational Governance

DGX agent

arXiv:2604.09620v1 Announce Type: cross Abstract: Large language models are increasingly used to support organizational decisions from hiring to governance, raising fairness concerns in AI-assisted ev

safetyarxiv-cs-ai
14 Apr 2026
Local Ai

LTX 2.3 Lora Training - Data Set Captioning

DGX agent

This Reddit thread from r/StableDiffusion discusses best practices for captioning training datasets when fine-tuning LoRA models on LTX 2.3, Lightricks' video generation model. The community emphasis

local-air-stablediffusion
14 Apr 2026
Safety

MAESTRO: Meta-learning Adaptive Estimation of Scalarization Trade-offs for Reward Optimization

DGX agent

arXiv:2601.07208v2 Announce Type: replace-cross Abstract: Group-Relative Policy Optimization (GRPO) has emerged as an efficient paradigm for aligning Large Language Models (LLMs), yet its efficacy is

safetyarxiv-cs-cl
14 Apr 2026
Safety

MARLIN: Multi-Agent Reinforcement Learning Guided by Language-Based Inter-Robot Negotiation

DGX agent

arXiv:2410.14383v4 Announce Type: replace Abstract: Multi-agent reinforcement learning is a key method for training multi-robot systems. Through rewarding or punishing robots over a series of episodes

safetyarxiv-cs-ro
14 Apr 2026
Model Releases

MathAgent: Adversarial Evolution of Constraint Graphs for Mathematical Reasoning Data Synthesis

DGX agent

arXiv:2604.11188v1 Announce Type: cross Abstract: Synthesizing high-quality mathematical reasoning data without human priors remains a significant challenge. Current approaches typically rely on seed

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

MAVEN-T: Multi-Agent enVironment-aware Enhanced Neural Trajectory predictor with Reinforcement Learning

DGX agent

arXiv:2604.10169v1 Announce Type: new Abstract: Trajectory prediction remains a critical yet challenging component in autonomous driving systems, requiring sophisticated reasoning capabilities while m

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

MCERF: Advancing Multimodal LLM Evaluation of Engineering Documentation with Enhanced Retrieval

DGX agent

arXiv:2604.09552v1 Announce Type: cross Abstract: Engineering rulebooks and technical standards contain multimodal information like dense text, tables, and illustrations that are challenging for retri

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

MedLVR: Latent Visual Reasoning for Reliable Medical Visual Question Answering

DGX agent

arXiv:2604.09757v1 Announce Type: cross Abstract: Medical vision--language models (VLMs) have shown strong potential for medical visual question answering (VQA), yet their reasoning remains largely te

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Most AI assistants wait for you to ask. But a truly useful agent should notice you need help before you say anything. New research takes a s…

DGX agent

Most AI assistants wait for you to ask. But a truly useful agent should notice you need help before you say anything. New research takes a serious shot at building proactive agents that work in real t

model-releasesdair-ai--x
14 Apr 2026
Applications

Navigating the generative AI journey: The Path-to-Value framework from AWS

DGX agent

The AWS Generative AI Path-to-Value (P2V) framework is a structured mental model and practical guide designed to help organizations move generative AI initiatives from ideation and experimentation thr

applicationsaws-ml-blog
14 Apr 2026
Local Ai

Openclaw with Gemma4 26B extremely slow and forget stuff

DGX agent

Users in the r/ollama community report that running OpenClaw with the Gemma4 26B model via Ollama results in pathologically slow first-turn performance, with the degree of slowdown scaling with OpenCl

local-air-ollama
14 Apr 2026
Agents

Our technical report, including how we designed the reward and the performance/latency Pareto frontier, is live here: https://cognition.ai/b…

DGX agent

Cognition AI has published a technical report detailing the design methodology behind their reward system, likely related to their Devin AI software engineering agent. The report covers the performanc

agentscognition-ai--x
14 Apr 2026
Hardware

Parisians: we're running an open source AI art hackathon with LTX + NVIDIA this Saturday

DGX agent

A Reddit post on r/StableDiffusion announces an in-person open source AI art hackathon held in Paris, co-organized with LTX (Lightricks' open source AI video model) and NVIDIA. The event likely invite

hardwarer-stablediffusion
14 Apr 2026
Model Releases

Please Make it Sound like Human: Encoder-Decoder vs. Decoder-Only Transformers for AI-to-Human Text Style Transfer

DGX agent

arXiv:2604.11687v1 Announce Type: new Abstract: AI-generated text has become common in academic and professional writing, prompting research into detection methods. Less studied is the reverse: system

model-releasesarxiv-cs-cl
14 Apr 2026
Safety

PRISM Risk Signal Framework: Hierarchy-Based Red Lines for AI Behavioral Risk

DGX agent

arXiv:2604.11070v1 Announce Type: new Abstract: Current approaches to AI safety define red lines at the case level: specific prompts, specific outputs, specific harms. This paper argues that red lines

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Progressive Multimodal Interaction Network for Reliable Quantification of Fish Feeding Intensity in Aquaculture

DGX agent

arXiv:2506.14170v3 Announce Type: replace-cross Abstract: Accurate quantification of fish feeding intensity is crucial for precision feeding in aquaculture, as it directly affects feed utilization and

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Quantization Dominates Rank Reduction for KV-Cache Compression

DGX agent

arXiv:2604.11501v1 Announce Type: cross Abstract: We compare two strategies for compressing the KV cache in transformer inference: rank reduction (discard dimensions) and quantization (keep all dimens

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Relational Preference Encoding in Looped Transformer Internal States

DGX agent

arXiv:2604.09870v1 Announce Type: cross Abstract: We investigate how looped transformers encode human preference in their internal iteration states. Using Ouro-2.6B-Thinking, a 2.6B-parameter looped t

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Robust Adversarial Policy Optimization Under Dynamics Uncertainty

DGX agent

arXiv:2604.10974v1 Announce Type: new Abstract: Reinforcement learning (RL) policies often fail under dynamics that differ from training, a gap not fully addressed by domain randomization or existing

model-releasesarxiv-cs-lg
14 Apr 2026
Local Ai

Saar-Voice: A Multi-Speaker Saarbrucken Dialect Speech Corpus

DGX agent

arXiv:2604.11803v1 Announce Type: new Abstract: Natural language processing (NLP) and speech technologies have made significant progress in recent years; however, they remain largely focused on standa

local-aiarxiv-cs-cl
14 Apr 2026
Safety

See Fair, Speak Truth: Equitable Attention Improves Grounding and Reduces Hallucination in Vision-Language Alignment

DGX agent

arXiv:2604.09749v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) frequently hallucinate objects that are absent from the visual input, often because attention during decoding i

safetyarxiv-cs-cv
14 Apr 2026
Model Releases

Semantic Manipulation Localization

DGX agent

arXiv:2604.10132v1 Announce Type: cross Abstract: Image Manipulation Localization (IML) aims to identify edited regions in an image. However, with the increasing use of modern image editing and genera

model-releasesarxiv-cs-ai
14 Apr 2026
← Previous
1…637638639640641…1371
Next →