AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlog
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,490 results
Model Releases

RoboWM-Bench: A Benchmark for Evaluating World Models in Robotic Manipulation

DGX agent

arXiv:2604.19092v1 Announce Type: cross Abstract: Recent advances in large-scale video world models have enabled increasingly realistic future prediction, raising the prospect of leveraging imagined v

model-releasesarxiv-cs-ai
22 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

SafetyALFRED: Evaluating Safety-Conscious Planning of Multimodal Large Language Models

DGX agent

arXiv:2604.19638v1 Announce Type: new Abstract: Multimodal Large Language Models are increasingly adopted as autonomous agents in interactive environments, yet their ability to proactively address saf

model-releasesarxiv-cs-ai
22 Apr 2026
Local Ai

Self-Improving Tabular Language Models via Iterative Group Alignment

DGX agent

arXiv:2604.18966v1 Announce Type: cross Abstract: While language models have been adapted for tabular data generation, two fundamental limitations remain: (1) static fine-tuning produces models that c

local-aiarxiv-cs-ai
22 Apr 2026
Safety

TROJail: Trajectory-Level Optimization for Multi-Turn Large Language Model Jailbreaks with Process Rewards

DGX agent

arXiv:2512.07761v3 Announce Type: replace Abstract: Large language models have seen widespread adoption, yet they remain vulnerable to multi-turn jailbreak attacks, threatening their safe deployment.

safetyarxiv-cs-ai
22 Apr 2026
Model Releases

We've published new research on how we post-train models for accurate search-augmented answers. Our SFT + RL pipeline improves search, citat…

DGX agent

We've published new research on how we post-train models for accurate search-augmented answers. Our SFT + RL pipeline improves search, citation quality, instruction following, and efficiency. With Qwe

model-releasesperplexity--x
22 Apr 2026
Model Releases

A Benchmark Study of Segmentation Models and Adaptation Strategies for Landslide Detection from Satellite Imagery

DGX agent

arXiv:2604.16663v1 Announce Type: new Abstract: Landslide detection from high resolution satellite imagery is a critical task for disaster response and risk assessment, yet the relative effectiveness

model-releasesarxiv-cs-cv
21 Apr 2026
Applications

A multimodal and temporal foundation model for virtual patient representations at healthcare system scale

DGX agent

arXiv:2604.18570v1 Announce Type: cross Abstract: Modern medicine generates vast multimodal data across siloed systems, yet no existing model integrates the full breadth and temporal depth of the clin

applicationsarxiv-cs-cl
21 Apr 2026
Agents

A Rapid Deployment Pipeline for Autonomous Humanoid Grasping Based on Foundation Models

DGX agent

arXiv:2604.17258v1 Announce Type: new Abstract: Deploying a humanoid robot to manipulate a new object has traditionally required one to two days of effort: data collection, manual annotation, 3D model

agentsarxiv-cs-ro
21 Apr 2026
Model Releases

Adversarial Humanities Benchmark: Results on Stylistic Robustness in Frontier Model Safety

DGX agent

arXiv:2604.18487v1 Announce Type: new Abstract: The Adversarial Humanities Benchmark (AHB) evaluates whether model safety refusals survive a shift away from familiar harmful prompt forms. Starting fro

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

[AINews] Moonshot Kimi K2.6: the world's leading Open Model refreshes to catch up to Opus 4.6 (ahead of DeepSeek v4?)

DGX agent

Moonshot's Kimi K2.6 represents a significant update to their open-source language model, positioning it to compete with Anthropic's Claude Opus 4.6 and potentially ahead of DeepSeek v4. The refresh a

model-releaseslatent-space
21 Apr 2026
Safety

Aligning Language Models for Lyric-to-Melody Generation with Rule-Based Musical Constraints

DGX agent

arXiv:2604.18489v1 Announce Type: cross Abstract: Large Language Models (LLMs) show promise in lyric-to-melody generation, but models trained with Supervised Fine-Tuning (SFT) often produce musically

safetyarxiv-cs-cl
21 Apr 2026
Applications

Appearance-free Action Recognition: Zero-shot Generalization in Humans and a Two-Pathway Model

DGX agent

arXiv:2604.16675v1 Announce Type: new Abstract: Action recognition is a fundamental ability for social species. Yet, its underlying computations are not well understood. Classical psychophysical studi

applicationsarxiv-cs-cv
21 Apr 2026
Research

Applications of deep generative models to DNA reaction kinetics and to cryogenic electron microscopy

DGX agent

arXiv:2604.16851v1 Announce Type: cross Abstract: This dissertation explores how deep generative models can advance the analysis of challenging biological problems by integrating domain knowledge with

researcharxiv-cs-cv
21 Apr 2026
Safety

Audio-DeepThinker: Progressive Reasoning-Aware Reinforcement Learning for High-Quality Chain-of-Thought Emergence in Audio Language Models

DGX agent

arXiv:2604.18187v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) have made significant progress in audio understanding, yet they primarily operate as perception-and-answer systems

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Auto-encoder model for faster generation of effective one-body gravitational waveform approximations

DGX agent

arXiv:2511.12642v2 Announce Type: replace-cross Abstract: Upgrades to current gravitational wave detectors for the next observation run and upcoming third-generation observatories, like the Einstein t

model-releasesarxiv-cs-lg
21 Apr 2026
Research

Beyond the Failures: Rethinking Foundation Models in Pathology

DGX agent

arXiv:2510.23807v5 Announce Type: replace-cross Abstract: Despite their successes in vision and language, foundation models have stumbled in pathology, revealing low accuracy, instability, and heavy c

researcharxiv-cs-cv
21 Apr 2026
Model Releases

CaTS-Bench: Can Language Models Describe Time Series?

DGX agent

arXiv:2509.20823v5 Announce Type: replace-cross Abstract: Time series captioning, the task of describing time series in natural language, requires numeric and temporal reasoning, trend interpretation,

model-releasesarxiv-cs-cv
21 Apr 2026
Applications

Clarifai says it has deleted 3M OkCupid user photos and facial-recognition models trained on them after the US FTC settled with OkCupid over privacy violations (Jody Godoy/Reuters)

DGX agent

Jody Godoy / Reuters: Clarifai says it has deleted 3M OkCupid user photos and facial-recognition models trained on them after the US FTC settled with OkCupid over privacy violations — Artificial intel

applicationstechmeme
21 Apr 2026
Model Releases

DifFoundMAD: Foundation Models meet Differential Morphing Attack Detection

DGX agent

arXiv:2604.17961v1 Announce Type: new Abstract: In this work, we introduce DifFoundMAD, a parameter-efficient D-MAD framework that exploits the generalisation capabilities of vision foundation models

model-releasesarxiv-cs-cv
21 Apr 2026
Tutorials

Dissipative Latent Residual Physics-Informed Neural Networks for Modeling and Identification of Electromechanical Systems

DGX agent

arXiv:2604.18277v1 Announce Type: new Abstract: Accurate dynamical modeling is essential for simulation and control of embodied systems, yet first-principles models of electromechanical systems often

tutorialsarxiv-cs-lg
21 Apr 2026
Research

Does AI See like Art Historians? Interpreting How Vision Language Models Recognize Artistic Style

DGX agent

arXiv:2603.11024v2 Announce Type: replace Abstract: VLMs have become increasingly proficient at a range of computer vision tasks, such as visual question answering and object detection. This includes

researcharxiv-cs-cv
21 Apr 2026
Research

Dual-End Consistency Model

DGX agent

arXiv:2602.10764v2 Announce Type: replace Abstract: The slow iterative sampling nature remains a major bottleneck for the practical deployment of diffusion and flow-based generative models. While cons

researcharxiv-cs-cv
21 Apr 2026
Model Releases

Efficient Task Adaptation in Large Language Models via Selective Parameter Optimization

DGX agent

arXiv:2604.17051v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated excellent performance in general language understanding, generation and other tasks. However, when fine-t

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Embedding Arithmetic: A Lightweight, Tuning-Free Framework for Post-hoc Bias Mitigation in Text-to-Image Models

DGX agent

arXiv:2604.18167v1 Announce Type: new Abstract: Modern text-to-image (T2I) models amplify harmful societal biases, challenging their ethical deployment. We introduce an inference-time method that reli

model-releasesarxiv-cs-cv
21 Apr 2026
Research

Enhancing Trust in Large Language Models via Uncertainty-Calibrated Fine-Tuning

DGX agent

arXiv:2412.02904v2 Announce Type: replace Abstract: Large language models (LLMs) have revolutionized the field of natural language processing with their impressive reasoning and question-answering cap

researcharxiv-cs-cl
21 Apr 2026
Research

Geometry-Guided 3D Visual Token Pruning for Video-Language Models

DGX agent

arXiv:2604.18260v1 Announce Type: new Abstract: Multimodal large language models have demonstrated remarkable capabilities in 2D vision, motivating their extension to 3D scene understanding. Recent st

researcharxiv-cs-cv
21 Apr 2026
Applications

How Training Data Shapes the Use of Parametric and In-Context Knowledge in Language Models

DGX agent

arXiv:2510.02370v3 Announce Type: replace Abstract: Large language models leverage both parametric knowledge acquired during pretraining and in-context knowledge provided at inference time. Crucially,

applicationsarxiv-cs-cl
21 Apr 2026
Agents

Human Cognition in Machines: A Unified Perspective of World Models

DGX agent

arXiv:2604.16592v1 Announce Type: cross Abstract: This comprehensive report distinguishes prior works by the cognitive functions they innovate. Many works claim an almost 'human-like' cognitive capabi

agentsarxiv-cs-cv
21 Apr 2026
Model Releases

I came up with a somewhat foolish new benchmark for testing image generation models, to exercise the new ChatGPT Images 2.0: 'Do a where's W…

DGX agent

I came up with a somewhat foolish new benchmark for testing image generation models, to exercise the new ChatGPT Images 2.0: 'Do a where's Waldo style image but it's where is the raccoon holding a ham

model-releasessimon-willison--x
21 Apr 2026
Model Releases

I find that open weights models over-perform on benchmarks compared to actual real-world usage, and Kimi feels like no exception. For exampl…

DGX agent

I find that open weights models over-perform on benchmarks compared to actual real-world usage, and Kimi feels like no exception. For example, a small amount of use will show that Kimi is not as good

model-releasesethan-mollick--x
21 Apr 2026
Model Releases

ICAT: Incident-Case-Grounded Adaptive Testing for Physical-Risk Prediction in Embodied World Models

DGX agent

arXiv:2604.16405v1 Announce Type: cross Abstract: Video-generative world models are increasingly used as neural simulators for embodied planning and policy learning, yet their ability to predict physi

model-releasesarxiv-cs-cv
21 Apr 2026
Safety

Infrastructure-Centric World Models: Bridging Temporal Depth and Spatial Breadth for Roadside Perception

DGX agent

arXiv:2604.17651v1 Announce Type: new Abstract: World models, generative AI systems that simulate how environments evolve, are transforming autonomous driving, yet all existing approaches adopt an ego

safetyarxiv-cs-cv
21 Apr 2026
Model Releases

Kimi K2.6 @Kimi_Moonshot is the new leading open-weights agent model, landing at #4 on Claw-Eval (Pass^3: 62.3%). Key takeaways: - 👑 Best o…

DGX agent

Kimi K2.6 @Kimi_Moonshot is the new leading open-weights agent model, landing at #4 on Claw-Eval (Pass^3: 62.3%). Key takeaways: - 👑 Best open-source agent, period: Pass^3 of 62.3% is the highest of a

model-releaseskimi-moonshot--x
21 Apr 2026
Model Releases

Lizard: An Efficient Linearization Framework for Large Language Models

DGX agent

arXiv:2507.09025v4 Announce Type: replace Abstract: We propose Lizard, a linearization framework that transforms pretrained Transformer-based Large Language Models (LLMs) into subquadratic architectur

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MetaLint: Easy-to-Hard Generalization for Code Linting

DGX agent

arXiv:2507.11687v4 Announce Type: replace-cross Abstract: Large language models excel at code generation but struggle with code linting, particularly in generalizing to unseen or evolving best practic

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Neural Network-Based Score Estimation in Diffusion Models: Optimization and Generalization

DGX agent

arXiv:2401.15604v4 Announce Type: replace Abstract: Diffusion models have become a leading paradigm in generative AI, with score estimation via denoising score matching as a central component. While r

researcharxiv-cs-lg
21 Apr 2026
Research

Ouroboros: Single-step Diffusion Models for Cycle-consistent Forward and Inverse Rendering

DGX agent

arXiv:2508.14461v3 Announce Type: replace Abstract: While multi-step diffusion models have advanced both forward and inverse rendering, existing approaches often treat these problems independently, le

researcharxiv-cs-cv
21 Apr 2026
Model Releases

Please refuse to answer me! Mitigating Over-Refusal in Large Language Models via Adaptive Contrastive Decoding

DGX agent

arXiv:2604.17132v1 Announce Type: new Abstract: Safety-aligned large language models (LLMs) often generate refusal responses to harmless queries due to the over-refusal problem. However, existing meth

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

PoliLegalLM: A Technical Report on a Large Language Model for Political and Legal Affairs

DGX agent

arXiv:2604.17543v1 Announce Type: new Abstract: Large language models (LLMs) have achieved remarkable success in general-domain tasks, yet their direct application to the legal domain remains challeng

safetyarxiv-cs-cl
21 Apr 2026
Research

Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning

DGX agent

arXiv:2502.02871v2 Announce Type: replace Abstract: Scientific reasoning, the process through which humans apply logic, evidence, and critical thinking to explore and interpret scientific phenomena, i

researcharxiv-cs-cl
21 Apr 2026
Research

Process Reward Models Meet Planning: Generating Precise and Scalable Datasets for Step-Level Rewards

DGX agent

arXiv:2604.17957v1 Announce Type: new Abstract: Process Reward Models (PRMs) have emerged as a powerful tool for providing step-level feedback when evaluating the reasoning of Large Language Models (L

researcharxiv-cs-cl
21 Apr 2026
Model Releases

ProfVLM: A lightweight video-language model for multi-view proficiency estimation

DGX agent

arXiv:2509.26278v4 Announce Type: replace-cross Abstract: Most existing approaches formulate action quality assessment and skill proficiency estimation as discriminative prediction tasks, typically pr

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Real-Time Visual Attribution Streaming in Thinking Model

DGX agent

arXiv:2604.16587v1 Announce Type: new Abstract: We present an amortized framework for real-time visual attribution streaming in multimodal thinking models. When these models generate code from a scree

researcharxiv-cs-cv
21 Apr 2026
Research

REALM: Reliable Expertise-Aware Language Model Fine-Tuning from Noisy Annotations

DGX agent

arXiv:2604.17289v1 Announce Type: new Abstract: Supervised fine-tuning of large language models relies on human-annotated data, yet annotation pipelines routinely involve multiple crowdworkers of hete

researcharxiv-cs-lg
21 Apr 2026
Safety

RemoteShield: Enable Robust Multimodal Large Language Models for Earth Observation

DGX agent

arXiv:2604.17243v1 Announce Type: new Abstract: A robust Multimodal Large Language Model (MLLM) for Earth Observation should maintain consistent interpretation and reasoning under realistic input vari

safetyarxiv-cs-cv
21 Apr 2026
Model Releases

ReTraceQA: Evaluating Reasoning Traces of Small Language Models in Commonsense Question Answering

DGX agent

arXiv:2510.09351v2 Announce Type: replace Abstract: While Small Language Models (SLMs) have demonstrated promising performance on an increasingly wide array of commonsense reasoning benchmarks, curren

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

S-GRPO: Unified Post-Training for Large Vision-Language Models

DGX agent

arXiv:2604.16557v1 Announce Type: cross Abstract: Current post-training methodologies for adapting Large Vision-Language Models (LVLMs) generally fall into two paradigms: Supervised Fine-Tuning (SFT)

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Spatiotemporal Sycophancy: Negation-Based Gaslighting in Video Large Language Models

DGX agent

arXiv:2604.17873v1 Announce Type: new Abstract: Video Large Language Models (Vid-LLMs) have demonstrated remarkable performance in video understanding tasks, yet their robustness under conversational

model-releasesarxiv-cs-cv
21 Apr 2026
← Previous
1…9192939495…1261
Next →