AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,520 results
1 May 2026

EDU-CIRCUIT-HW: Evaluating Multimodal Large Language Models on Real-World University-Level STEM Student Handwritten Solutions

Model ReleasesDGX agent

arXiv:2602.00095v3 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) hold significant promise for revolutionizing traditional education and reducing teachers' workload. H

Entropy of Ukrainian

Model ReleasesDGX agent

arXiv:2604.27534v1 Announce Type: new Abstract: In natural language processing, the entropy of a language is a measure of its unpredictability and complexity. The first study on this subject was condu

Event with @googlegemma next week! Hang out with the fellow builders and members of the Gemma + LM teams at LM Studio HQ in NYC. When: Monda…

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Event with @googlegemma next week! Hang out with the fellow builders and members of the Gemma + LM teams at LM Studio HQ in NYC. When: Monday, May 11th RSVP: required, link below High likelihood of pi

Exploring Interaction Paradigms for LLM Agents in Scientific Visualization

Model ReleasesDGX agent

arXiv:2604.27996v1 Announce Type: new Abstract: This paper examines how different types of large language model (LLM) agents perform on scientific visualization (SciVis) tasks, where users generate vi

Exploring the Limits of Pruning: Task-Specific Neurons, Model Collapse, and Recovery in Task-Specific Large Language Models

Model ReleasesDGX agent

arXiv:2604.27115v1 Announce Type: new Abstract: Neuron pruning is widely used to reduce the computational cost and parameter footprint of large language models, yet it remains unclear whether neurons

Fake3DGS: A Benchmark for 3D Manipulation Detection in Neural Rendering

Model ReleasesDGX agent

arXiv:2604.27590v1 Announce Type: new Abstract: Recent advances in 3D reconstruction and neural rendering,particularly 3D Gaussian Splatting, make it feasible and simple to edit 3D scenes and re-rende

Federated Knowledge Distillation for Multi-Model Architectures Lithography Hotspot Detection

Model ReleasesDGX agent

arXiv:2501.04066v2 Announce Type: replace Abstract: As a special type of multimedia data, Lithography Hotspot Detection (LHD) training often requires stronger privacy protection than conventional mult

Fidelity, Diversity, and Privacy: A Multi-Dimensional LLM Evaluation for Clinical Data Augmentation

Model ReleasesDGX agent

arXiv:2604.27014v1 Announce Type: new Abstract: The scarcity of high-quality annotated medical data, particularly in mental health, poses a significant bottleneck for training robust machine learning

FinChain: A Symbolic Benchmark for Verifiable Chain-of-Thought Financial Reasoning

Model ReleasesDGX agent

arXiv:2506.02515v4 Announce Type: replace-cross Abstract: Multi-step symbolic reasoning is essential for robust financial analysis; yet, current benchmarks largely overlook this capability. Existing d

FineState-Bench: Benchmarking State-Conditioned Grounding for Fine-grained GUI State Setting

Model ReleasesDGX agent

arXiv:2604.27974v1 Announce Type: new Abstract: Despite the rapid progress of large vision-language models (LVLMs), fine-grained, state-conditioned GUI interaction remains challenging. Current evaluat

Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs

Model ReleasesDGX agent

arXiv:2506.07180v3 Announce Type: replace-cross Abstract: As video large language models (Video-LLMs) become increasingly integrated into real-world applications that demand grounded multimodal reason

FluxMoE: Decoupling Expert Residency for High-Performance MoE Serving

Model ReleasesDGX agent

arXiv:2604.02715v2 Announce Type: replace Abstract: Mixture-of-Experts (MoE) models have become a dominant paradigm for scaling large language models, but their rapidly growing parameter sizes introdu

FreeOcc: Training-Free Embodied Open-Vocabulary Occupancy Prediction

Model ReleasesDGX agent

arXiv:2604.28115v1 Announce Type: cross Abstract: Existing learning-based occupancy prediction methods rely on large-scale 3D annotations and generalize poorly across environments. We present FreeOcc,

From Mirage to Grounding: Towards Reliable Multimodal Circuit-to-Verilog Code Generation

Model ReleasesDGX agent

arXiv:2604.27969v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) are increasingly used to translate visual artifacts into code, from UI mockups into HTML to scientific plots

From Test-taking to Cognitive Scaffolding: A Pedagogical Diagnostic Benchmark for LLMs on English Standardized Tests

Model ReleasesDGX agent

arXiv:2505.17056v2 Announce Type: replace-cross Abstract: As large language models (LLMs) are increasingly integrated into educational tools, current evaluations on standardized tests predominantly fo

From Unstructured Recall to Schema-Grounded Memory: Reliable AI Memory via Iterative, Schema-Aware Extraction

Model ReleasesDGX agent

arXiv:2604.27906v1 Announce Type: new Abstract: Persistent AI memory is often reduced to a retrieval problem: store prior interactions as text, embed them, and ask the model to recover relevant contex

Function-based Parametric Co-Design Optimization of Dexterous Hands

Model ReleasesDGX agent

arXiv:2604.27557v1 Announce Type: new Abstract: Despite advances in dexterous hand manipulation, robotic hand design is still largely decoupled from task-driven evaluation and control, limiting system

Gait Recognition via Deep Residual Networks and Multi-Branch Feature Fusion

Model ReleasesDGX agent

arXiv:2604.27353v1 Announce Type: new Abstract: Gait recognition has emerged as a compelling biometric modality for surveillance and security applications, offering inherent advantages such as non-int

Generalizable Sparse-View 3D Reconstruction from Unconstrained Images

Model ReleasesDGX agent

arXiv:2604.28193v1 Announce Type: new Abstract: Reconstructing 3D scenes from sparse, unposed images remains challenging under real-world conditions with varying illumination and transient occlusions.

Generalizing the Geometry of Model Merging Through Frechet Averages

Model ReleasesDGX agent

arXiv:2604.27155v1 Announce Type: new Abstract: Model merging aims to combine multiple models into one without additional training. Naive parameter-space averaging can be fragile under architectural s

Generate Your Talking Avatar from Video Reference

Model ReleasesDGX agent

arXiv:2604.27918v1 Announce Type: new Abstract: Existing talking avatar methods typically adopt an image-to-video pipeline conditioned on a static reference image within the same scene as the target g

Global Optimality for Constrained Exploration via Penalty Regularization

Model ReleasesDGX agent

arXiv:2604.28144v1 Announce Type: new Abstract: Efficient exploration is a central problem in reinforcement learning and is often formalized as maximizing the entropy of the state-action occupancy mea

GlowQ: Group-Shared LOw-Rank Approximation for Quantized LLMs

Model ReleasesDGX agent

arXiv:2603.25385v2 Announce Type: replace-cross Abstract: Quantization techniques such as BitsAndBytes, AWQ, and GPTQ are widely used as a standard method in deploying large language models but often

@GoogleAIStudio And this sample project was created on Canvas in @GeminiApp. It’s a high-speed rhythm game where you tap to the beat and col…

Model ReleasesDGX agent

@GoogleAIStudio And this sample project was created on Canvas in @GeminiApp. It’s a high-speed rhythm game where you tap to the beat and collect power-ups to remix the track. Watch as the numbers appe

@GoogleAIStudio @GeminiApp We can’t wait to see where your creativity takes you. Vibe code your countdown idea in @GoogleAIStudio or Canvas …

Model ReleasesDGX agent

@GoogleAIStudio @GeminiApp We can’t wait to see where your creativity takes you. Vibe code your countdown idea in @GoogleAIStudio or Canvas in @GeminiApp, then submit it here: http://goo.gle/codetheco

Grounding Agent Memory in Contextual Intent

Model ReleasesDGX agent

arXiv:2601.10702v2 Announce Type: replace-cross Abstract: Deploying large language models in long-horizon, goal-oriented interactions remains challenging because similar entities and facts recur under

GuideDog: A Real-World Egocentric Multimodal Dataset for Blind and Low-Vision Accessibility-Aware Guidance

Model ReleasesDGX agent

arXiv:2503.12844v2 Announce Type: replace Abstract: For people affected by blindness and low vision (BLV), safe and independent navigation remains a major challenge, impacting over 2.2 billion individ

HealthBench Professional: Evaluating Large Language Models on Real Clinician Chats

Model ReleasesDGX agent

arXiv:2604.27470v1 Announce Type: new Abstract: Millions of clinicians use ChatGPT to support clinical care, but evaluations of the most common use cases in model-clinician conversations are limited.

HERMES++: Toward a Unified Driving World Model for 3D Scene Understanding and Generation

Model ReleasesDGX agent

arXiv:2604.28196v1 Announce Type: new Abstract: Driving world models serve as a pivotal technology for autonomous driving by simulating environmental dynamics. However, existing approaches predominant

Hi Singapore 🇸🇬, meet Codex 🩵 With Codex, ANYONE can build and create. We’re turning that energy up this May. We’re a diamond sponsor 💎 …

Model ReleasesDGX agent

Hi Singapore 🇸🇬, meet Codex 🩵 With Codex, ANYONE can build and create. We’re turning that energy up this May. We’re a diamond sponsor 💎 at @aiDotEngineer Singapore, at a bunch of events, and hosting s

HighFM: Towards a Foundation Model for Learning Representations from High-Frequency Earth Observation Data

Model ReleasesDGX agent

arXiv:2604.04306v2 Announce Type: replace-cross Abstract: The increasing frequency and severity of climate related disasters have intensified the need for real time monitoring, early warning, and info

How Generative AI Disrupts Search: An Empirical Study of Google Search, Gemini, and AI Overviews

Model ReleasesDGX agent

arXiv:2604.27790v1 Announce Type: cross Abstract: Generative AI is being increasingly integrated into web search for the convenience it provides users. In this work, we aim to understand how generativ

HQ-UNet: A Hybrid Quantum-Classical U-Net with a Quantum Bottleneck for Remote Sensing Image Segmentation

Model ReleasesDGX agent

arXiv:2604.27206v1 Announce Type: new Abstract: Semantic segmentation in remote sensing is commonly addressed using classical deep learning architectures such as U-Net, which require a large number of

I have been testing DeepSeek-V4-Pro with the Pi coding agent. I am mindblown by how well it works out of the box. A few notes: I spent a few…

Model ReleasesDGX agent

I have been testing DeepSeek-V4-Pro with the Pi coding agent. I am mindblown by how well it works out of the box. A few notes: I spent a few hours building an LLM wiki with an agent powered entirely b

Improving Graph Few-shot Learning with Hyperbolic Space and Denoising Diffusion

Model ReleasesDGX agent

arXiv:2604.27462v1 Announce Type: cross Abstract: Graph few-shot learning, which focuses on effectively learning from only a small number of labeled nodes to quickly adapt to new tasks, has garnered s

iNaturalist Sightings

Model ReleasesDGX agent

Tool: iNaturalist Sightings I wanted to see my iNaturalist observations - across two separate accounts - grouped by when they occurred. I'm camping this weekend so I built this entirely on my phone us

Instruction Complexity Induces Positional Collapse in Adversarial LLM Evaluation

Model ReleasesDGX agent

arXiv:2604.27249v1 Announce Type: cross Abstract: When instructed to underperform on multiple-choice evaluations, do language models engage with question content or fall back on positional shortcuts?

Intent2Tx: Benchmarking LLMs for Translating Natural Language Intents into Ethereum Transactions

Model ReleasesDGX agent

arXiv:2604.27763v1 Announce Type: new Abstract: The emergence of Large Language Models (LLMs) offers a transformative interface for Web3, yet existing benchmarks fail to capture the complexity of tran

InteractWeb-Bench: Can Multimodal Agent Escape Blind Execution in Interactive Website Generation?

Model ReleasesDGX agent

arXiv:2604.27419v1 Announce Type: new Abstract: With the advancement of multimodal large language models (MLLMs) and coding agents, the website development has shifted from manual programming to agent

Iterative Definition Refinement for Zero-Shot Classification via LLM-Based Semantic Prototype Optimization

Model ReleasesDGX agent

arXiv:2604.27335v1 Announce Type: new Abstract: Web filtering systems rely on accurate web content classification to block cyber threats, prevent data exfiltration, and ensure compliance. However, cla

Iterative Multimodal Retrieval-Augmented Generation for Medical Question Answering

Model ReleasesDGX agent

arXiv:2604.27724v1 Announce Type: new Abstract: Medical retrieval-augmented generation (RAG) systems typically operate on text chunks extracted from biomedical literature, discarding the rich visual c

ITS-Mina: A Harris Hawks Optimization-Based All-MLP Framework with Iterative Refinement and External Attention for Multivariate Time Series Forecasting

Model ReleasesDGX agent

arXiv:2604.27981v1 Announce Type: cross Abstract: Multivariate time series forecasting plays a pivotal role in numerous real-world applications, including financial analysis, energy management, and tr

JI-ADF: Joint-Individual Learning with Adaptive Decision Fusion for Multimodal Skin Lesion Classification

Model ReleasesDGX agent

arXiv:2604.27343v1 Announce Type: new Abstract: Skin lesion classification is essential for early dermatological diagnosis, yet many existing computer-aided systems rely primarily on dermoscopic image

Judge, Then Drive: A Critic-Centric Vision Language Action Framework for Autonomous Driving

Model ReleasesDGX agent

arXiv:2604.27366v1 Announce Type: new Abstract: Recent advances in vision language action (VLA) models have shown remarkable potential for autonomous driving by directly mapping multimodal inputs to c

K2MUSE: A human lower-limb multimodal walking dataset spanning task and acquisition variability for rehabilitation robotics

Model ReleasesDGX agent

arXiv:2504.14602v2 Announce Type: replace-cross Abstract: The natural interaction and control performance of lower limb rehabilitation robots are closely linked to biomechanical information from vario

KellyBench: A Benchmark for Long-Horizon Sequential Decision Making

Model ReleasesDGX agent

arXiv:2604.27865v1 Announce Type: new Abstract: Language models are saturating benchmarks for procedural tasks with narrow objectives. But they are increasingly being deployed in long-horizon, non-sta

Language Models Refine Mechanical Linkage Designs Through Symbolic Reflection and Modular Optimisation

Model ReleasesDGX agent

arXiv:2604.27962v1 Announce Type: new Abstract: Designing mechanical linkages involves combinatorial topology selection and continuous parameter fitting. We show that language models can systematicall

LaST-R1: Reinforcing Action via Adaptive Physical Latent Reasoning for VLA Models

Model ReleasesDGX agent

arXiv:2604.28192v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have increasingly incorporated reasoning mechanisms for complex robotic manipulation. However, existing approaches

Learning Generalizable Multimodal Representations for Software Vulnerability Detection

Model ReleasesDGX agent

arXiv:2604.25711v2 Announce Type: replace-cross Abstract: Source code and its accompanying comments are complementary yet naturally aligned modalities-code encodes structural logic while comments capt

Learning Rate Engineering: From Coarse Single Parameter to Layered Evolution

Model ReleasesDGX agent

arXiv:2604.27295v1 Announce Type: new Abstract: Learning rate scheduling has evolved from the single global fixed rate of early SGD to sophisticated layer-wise adaptive strategies. We systematize this

Learning to Forget: Continual Learning with Adaptive Weight Decay

Model ReleasesDGX agent

arXiv:2604.27063v1 Announce Type: new Abstract: Continual learning agents with finite capacity must balance acquiring new knowledge with retaining the old. This requires controlled forgetting of knowl

Lightweight Distillation of SAM 3 and DINOv3 for Edge-Deployable Individual-Level Livestock Monitoring and Longitudinal Visual Analytics

Model ReleasesDGX agent

arXiv:2604.27128v1 Announce Type: cross Abstract: Foundation-model pipelines for individual-level livestock monitoring -- combining open-vocabulary detection, promptable video segmentation, and self-s

LLM-Guided Runtime Parameter Optimization for Energy-Efficient Model Inference

Model ReleasesDGX agent

arXiv:2604.27032v1 Announce Type: cross Abstract: Large Language Models (LLMs) have become an integral part of many real-world workflows. However, LLMs consume a lot of energy, which becomes a large c

Local AI is about to be competitive. I will do everything in my power to make it better than Claude desktop/claude code by end of year.

Model ReleasesDGX agent

Clem Delangue expresses commitment to advancing local AI models to compete with Anthropic's Claude Desktop and Claude Code offerings by year-end. The statement suggests focus on improving local AI cap

Lost in Space? Vision-Language Models Struggle with Relative Camera Pose Estimation

Model ReleasesDGX agent

arXiv:2601.22228v2 Announce Type: replace-cross Abstract: We study whether vision-language models (VLMs) can solve relative camera pose estimation (RCPE) from image pairs, a direct test of multi-view

Low Rank Adaptation for Adversarial Perturbation

Model ReleasesDGX agent

arXiv:2604.27487v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA), which leverages the insight that model updates typically reside in a low-dimensional space, has significantly improved the t

M-DaQ: Retrieving Samples with Multilingual Diversity and Quality for Instruction Fine-Tuning Datasets

Model ReleasesDGX agent

arXiv:2509.15549v2 Announce Type: replace Abstract: Multilingual instruction fine-tuning (IFT) empowers large language models to generalize across diverse linguistic and cultural contexts; however, hi

MAEO: Multiobjective Animorphic Ensemble Optimization for Scalable Large-scale Engineering Applications

Model ReleasesDGX agent

arXiv:2604.26973v1 Announce Type: cross Abstract: Multiobjective optimization remains challenging for many scientific and engineering problems due to the need to balance convergence, diversity, and co

Make sure to read the blog post for a detailed analysis of frontier model failure modes: https://arcprize.org/blog/arc-agi-3-gpt-5-5-opus-4-…

Model ReleasesDGX agent

Francois Chollet shared a blog post analyzing failure modes of frontier AI models, specifically examining performance on the ARC (Abstraction and Reasoning Corpus) AGI benchmark with models including

Mapping the Phase Diagram of the Vicsek Model with Machine Learning

Model ReleasesDGX agent

arXiv:2604.28167v1 Announce Type: cross Abstract: In this study, we use machine learning to classify and interpolate the phase structure of the Vicsek flocking model across the three-dimensional param

← Previous
1…294295296297298…376
Next →