AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,555 results
Model Releases

Grounding Agent Memory in Contextual Intent

DGX agent

arXiv:2601.10702v2 Announce Type: replace-cross Abstract: Deploying large language models in long-horizon, goal-oriented interactions remains challenging because similar entities and facts recur under

model-releasesarxiv-cs-ai
1 May 2026
Model Releases
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

GuideDog: A Real-World Egocentric Multimodal Dataset for Blind and Low-Vision Accessibility-Aware Guidance

DGX agent

arXiv:2503.12844v2 Announce Type: replace Abstract: For people affected by blindness and low vision (BLV), safe and independent navigation remains a major challenge, impacting over 2.2 billion individ

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

HealthBench Professional: Evaluating Large Language Models on Real Clinician Chats

DGX agent

arXiv:2604.27470v1 Announce Type: new Abstract: Millions of clinicians use ChatGPT to support clinical care, but evaluations of the most common use cases in model-clinician conversations are limited.

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

HERMES++: Toward a Unified Driving World Model for 3D Scene Understanding and Generation

DGX agent

arXiv:2604.28196v1 Announce Type: new Abstract: Driving world models serve as a pivotal technology for autonomous driving by simulating environmental dynamics. However, existing approaches predominant

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

Hi Singapore 🇸🇬, meet Codex 🩵 With Codex, ANYONE can build and create. We’re turning that energy up this May. We’re a diamond sponsor 💎 …

DGX agent

Hi Singapore 🇸🇬, meet Codex 🩵 With Codex, ANYONE can build and create. We’re turning that energy up this May. We’re a diamond sponsor 💎 at @aiDotEngineer Singapore, at a bunch of events, and hosting s

model-releasesswyx--x
1 May 2026
Model Releases

HighFM: Towards a Foundation Model for Learning Representations from High-Frequency Earth Observation Data

DGX agent

arXiv:2604.04306v2 Announce Type: replace-cross Abstract: The increasing frequency and severity of climate related disasters have intensified the need for real time monitoring, early warning, and info

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

How Generative AI Disrupts Search: An Empirical Study of Google Search, Gemini, and AI Overviews

DGX agent

arXiv:2604.27790v1 Announce Type: cross Abstract: Generative AI is being increasingly integrated into web search for the convenience it provides users. In this work, we aim to understand how generativ

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

HQ-UNet: A Hybrid Quantum-Classical U-Net with a Quantum Bottleneck for Remote Sensing Image Segmentation

DGX agent

arXiv:2604.27206v1 Announce Type: new Abstract: Semantic segmentation in remote sensing is commonly addressed using classical deep learning architectures such as U-Net, which require a large number of

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

I have been testing DeepSeek-V4-Pro with the Pi coding agent. I am mindblown by how well it works out of the box. A few notes: I spent a few…

DGX agent

I have been testing DeepSeek-V4-Pro with the Pi coding agent. I am mindblown by how well it works out of the box. A few notes: I spent a few hours building an LLM wiki with an agent powered entirely b

model-releasesdair-ai--x
1 May 2026
Model Releases

Improving Graph Few-shot Learning with Hyperbolic Space and Denoising Diffusion

DGX agent

arXiv:2604.27462v1 Announce Type: cross Abstract: Graph few-shot learning, which focuses on effectively learning from only a small number of labeled nodes to quickly adapt to new tasks, has garnered s

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

iNaturalist Sightings

DGX agent

Tool: iNaturalist Sightings I wanted to see my iNaturalist observations - across two separate accounts - grouped by when they occurred. I'm camping this weekend so I built this entirely on my phone us

model-releasessimon-willison
1 May 2026
Model Releases

Instruction Complexity Induces Positional Collapse in Adversarial LLM Evaluation

DGX agent

arXiv:2604.27249v1 Announce Type: cross Abstract: When instructed to underperform on multiple-choice evaluations, do language models engage with question content or fall back on positional shortcuts?

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Intent2Tx: Benchmarking LLMs for Translating Natural Language Intents into Ethereum Transactions

DGX agent

arXiv:2604.27763v1 Announce Type: new Abstract: The emergence of Large Language Models (LLMs) offers a transformative interface for Web3, yet existing benchmarks fail to capture the complexity of tran

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

InteractWeb-Bench: Can Multimodal Agent Escape Blind Execution in Interactive Website Generation?

DGX agent

arXiv:2604.27419v1 Announce Type: new Abstract: With the advancement of multimodal large language models (MLLMs) and coding agents, the website development has shifted from manual programming to agent

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Iterative Definition Refinement for Zero-Shot Classification via LLM-Based Semantic Prototype Optimization

DGX agent

arXiv:2604.27335v1 Announce Type: new Abstract: Web filtering systems rely on accurate web content classification to block cyber threats, prevent data exfiltration, and ensure compliance. However, cla

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

Iterative Multimodal Retrieval-Augmented Generation for Medical Question Answering

DGX agent

arXiv:2604.27724v1 Announce Type: new Abstract: Medical retrieval-augmented generation (RAG) systems typically operate on text chunks extracted from biomedical literature, discarding the rich visual c

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

ITS-Mina: A Harris Hawks Optimization-Based All-MLP Framework with Iterative Refinement and External Attention for Multivariate Time Series Forecasting

DGX agent

arXiv:2604.27981v1 Announce Type: cross Abstract: Multivariate time series forecasting plays a pivotal role in numerous real-world applications, including financial analysis, energy management, and tr

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

JI-ADF: Joint-Individual Learning with Adaptive Decision Fusion for Multimodal Skin Lesion Classification

DGX agent

arXiv:2604.27343v1 Announce Type: new Abstract: Skin lesion classification is essential for early dermatological diagnosis, yet many existing computer-aided systems rely primarily on dermoscopic image

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

Judge, Then Drive: A Critic-Centric Vision Language Action Framework for Autonomous Driving

DGX agent

arXiv:2604.27366v1 Announce Type: new Abstract: Recent advances in vision language action (VLA) models have shown remarkable potential for autonomous driving by directly mapping multimodal inputs to c

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

K2MUSE: A human lower-limb multimodal walking dataset spanning task and acquisition variability for rehabilitation robotics

DGX agent

arXiv:2504.14602v2 Announce Type: replace-cross Abstract: The natural interaction and control performance of lower limb rehabilitation robots are closely linked to biomechanical information from vario

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

KellyBench: A Benchmark for Long-Horizon Sequential Decision Making

DGX agent

arXiv:2604.27865v1 Announce Type: new Abstract: Language models are saturating benchmarks for procedural tasks with narrow objectives. But they are increasingly being deployed in long-horizon, non-sta

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Language Models Refine Mechanical Linkage Designs Through Symbolic Reflection and Modular Optimisation

DGX agent

arXiv:2604.27962v1 Announce Type: new Abstract: Designing mechanical linkages involves combinatorial topology selection and continuous parameter fitting. We show that language models can systematicall

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

LaST-R1: Reinforcing Action via Adaptive Physical Latent Reasoning for VLA Models

DGX agent

arXiv:2604.28192v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have increasingly incorporated reasoning mechanisms for complex robotic manipulation. However, existing approaches

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

Learning Generalizable Multimodal Representations for Software Vulnerability Detection

DGX agent

arXiv:2604.25711v2 Announce Type: replace-cross Abstract: Source code and its accompanying comments are complementary yet naturally aligned modalities-code encodes structural logic while comments capt

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Learning Rate Engineering: From Coarse Single Parameter to Layered Evolution

DGX agent

arXiv:2604.27295v1 Announce Type: new Abstract: Learning rate scheduling has evolved from the single global fixed rate of early SGD to sophisticated layer-wise adaptive strategies. We systematize this

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Learning to Forget: Continual Learning with Adaptive Weight Decay

DGX agent

arXiv:2604.27063v1 Announce Type: new Abstract: Continual learning agents with finite capacity must balance acquiring new knowledge with retaining the old. This requires controlled forgetting of knowl

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

Lightweight Distillation of SAM 3 and DINOv3 for Edge-Deployable Individual-Level Livestock Monitoring and Longitudinal Visual Analytics

DGX agent

arXiv:2604.27128v1 Announce Type: cross Abstract: Foundation-model pipelines for individual-level livestock monitoring -- combining open-vocabulary detection, promptable video segmentation, and self-s

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

LLM-Guided Runtime Parameter Optimization for Energy-Efficient Model Inference

DGX agent

arXiv:2604.27032v1 Announce Type: cross Abstract: Large Language Models (LLMs) have become an integral part of many real-world workflows. However, LLMs consume a lot of energy, which becomes a large c

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

Local AI is about to be competitive. I will do everything in my power to make it better than Claude desktop/claude code by end of year.

DGX agent

Clem Delangue expresses commitment to advancing local AI models to compete with Anthropic's Claude Desktop and Claude Code offerings by year-end. The statement suggests focus on improving local AI cap

model-releasesclem-delangue--x
1 May 2026
Model Releases

Lost in Space? Vision-Language Models Struggle with Relative Camera Pose Estimation

DGX agent

arXiv:2601.22228v2 Announce Type: replace-cross Abstract: We study whether vision-language models (VLMs) can solve relative camera pose estimation (RCPE) from image pairs, a direct test of multi-view

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Low Rank Adaptation for Adversarial Perturbation

DGX agent

arXiv:2604.27487v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA), which leverages the insight that model updates typically reside in a low-dimensional space, has significantly improved the t

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

M-DaQ: Retrieving Samples with Multilingual Diversity and Quality for Instruction Fine-Tuning Datasets

DGX agent

arXiv:2509.15549v2 Announce Type: replace Abstract: Multilingual instruction fine-tuning (IFT) empowers large language models to generalize across diverse linguistic and cultural contexts; however, hi

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

MAEO: Multiobjective Animorphic Ensemble Optimization for Scalable Large-scale Engineering Applications

DGX agent

arXiv:2604.26973v1 Announce Type: cross Abstract: Multiobjective optimization remains challenging for many scientific and engineering problems due to the need to balance convergence, diversity, and co

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

Make sure to read the blog post for a detailed analysis of frontier model failure modes: https://arcprize.org/blog/arc-agi-3-gpt-5-5-opus-4-…

DGX agent

Francois Chollet shared a blog post analyzing failure modes of frontier AI models, specifically examining performance on the ARC (Abstraction and Reasoning Corpus) AGI benchmark with models including

model-releasesfrancois-chollet--x
1 May 2026
Model Releases

Mapping the Phase Diagram of the Vicsek Model with Machine Learning

DGX agent

arXiv:2604.28167v1 Announce Type: cross Abstract: In this study, we use machine learning to classify and interpolate the phase structure of the Vicsek flocking model across the three-dimensional param

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

Math Education Digital Shadows for facilitating learning with LLMs: Math performance, anxiety and confidence in simulated students and AIs

DGX agent

arXiv:2604.27618v1 Announce Type: new Abstract: To enhance LLMs' impact on math education, we need data on their mathematical prowess and biases across prompts. To fill this gap, we introduce MEDS (Ma

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

MCPHunt: An Evaluation Framework for Cross-Boundary Data Propagation in Multi-Server MCP Agents

DGX agent

arXiv:2604.27819v1 Announce Type: new Abstract: Multi-server MCP agents create an information-flow control problem: faithful tool composition can turn individually benign read/write permissions into c

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Measurement Risk in Supervised Financial NLP: Rubric and Metric Sensitivity on JF-ICR

DGX agent

arXiv:2604.27374v1 Announce Type: new Abstract: As LLMs become credible readers of earnings calls, investor-relations Q&A, guidance, and disclosure language, supervised financial NLP benchmarks increa

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

MiniCPM-o 4.5: Towards Real-Time Full-Duplex Omni-Modal Interaction

DGX agent

arXiv:2604.27393v1 Announce Type: new Abstract: Recent progress in multimodal large language models (MLLMs) has brought AI capabilities from static offline data processing to real-time streaming inter

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

Models Recall What They Violate: Constraint Adherence in Multi-Turn LLM Ideation

DGX agent

arXiv:2604.28031v1 Announce Type: new Abstract: When researchers iteratively refine ideas with large language models, do the models preserve fidelity to the original objective? We introduce DriftBench

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

Monitoring Neural Training with Topology: A Footprint-Predictable Collapse Index

DGX agent

arXiv:2604.26984v1 Announce Type: new Abstract: Representational collapse, where embeddings become anisotropic and lose multi-scale structure, can erode downstream performance long before performance

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

MultiBLiMP 1.0: A Massively Multilingual Benchmark of Linguistic Minimal Pairs

DGX agent

arXiv:2504.02768v4 Announce Type: replace Abstract: We introduce MultiBLiMP 1.0, a massively multilingual benchmark of linguistic minimal pairs, covering 101 languages and 2 types of subject-verb agre

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

NanoKnow: How to Know What Your Language Model Knows

DGX agent

arXiv:2602.20122v2 Announce Type: replace-cross Abstract: How do large language models (LLMs) know what they know? Answering this question has been difficult because pre-training data is often a 'blac

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

NashPG: A Policy Gradient Method with Iteratively Refined Regularization for Finding Nash Equilibria

DGX agent

arXiv:2510.18183v2 Announce Type: replace Abstract: Finding Nash equilibria in two-player zero-sum imperfect-information games remains a central challenge in multi-agent reinforcement learning. Recent

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

NeocorRAG: Less Irrelevant Information, More Explicit Evidence, and More Effective Recall via Evidence Chains

DGX agent

arXiv:2604.27852v1 Announce Type: cross Abstract: Although precise recall is a core objective in Retrieval-Augmented Generation (RAG), a critical oversight persists in the field: improvements in retri

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

ObjectGraph: From Document Injection to Knowledge Traversal -- A Native File Format for the Agentic Era

DGX agent

arXiv:2604.27820v1 Announce Type: new Abstract: Every document format in existence was designed for a human reader moving linearly through text. Autonomous LLM agents do not read - they retrieve. This

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

📢 Official Announcement: Qwen Partners with Fireworks AI to Accelerate Access to Qwen Family Models We are pleased to announce a strategic …

DGX agent

📢 Official Announcement: Qwen Partners with Fireworks AI to Accelerate Access to Qwen Family Models We are pleased to announce a strategic partnership between Qwen and Fireworks AI to deliver optimize

model-releasesqwen--x
1 May 2026
Model Releases

On 5/5 @realDanFu and team will discuss DSV4’s hybrid attention and KV cache efficiency, should be a great session!

DGX agent

On 5/5 @realDanFu and team will discuss DSV4’s hybrid attention and KV cache efficiency, should be a great session! Join us Tue 5/5: #DeepSeek-V4's hybrid attention + sparse MoE reduces KV cache up to

model-releasestogether-ai--x
1 May 2026
← Previous
1…368369370371372…470
Next →