AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,490 results
5 May 2026

ViM-Q: Scalable Algorithm-Hardware Co-Design for Vision Mamba Model Inference on FPGA

HardwareDGX agent

arXiv:2605.01935v1 Announce Type: cross Abstract: Vision Mamba (ViM) models offer a compelling efficiency advantage over Transformers by leveraging the linear complexity of State Space Models (SSMs),

what model are you choosing for coding tasks?

AgentsDGX agent

This post likely discusses Harrison Chase's preferred language model or AI system for handling coding tasks, potentially comparing different models' capabilities for programming work. As the creator o

4 May 2026

AirFM-DDA: Air-Interface Foundation Model in the Delay-Doppler-Angle Domain for AI-Native 6G

Local AiDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.00020v1 Announce Type: new Abstract: The success of large foundation models is catalyzing a new paradigm for AI-native 6G network design: wireless foundation models for physical layer desig

AlphaInventory: Evolving White-Box Inventory Policies via Large Language Models with Deployment Guarantees

Model ReleasesDGX agent

arXiv:2605.00369v1 Announce Type: new Abstract: We study how large language models can be used to evolve inventory policies in online, non-stationary environments. Our work is motivated by recent adva

An End-to-End Decision-Aware Multi-Scale Attention-Based Model for Explainable Autonomous Driving

AgentsDGX agent

arXiv:2605.00291v1 Announce Type: new Abstract: The application of computer vision is gradually increasing across various domains. They employ deep learning models with a black-box nature. Without the

At any point in time, you can safely resume to using Anthropic's models: ollama launch claude-desktop --restore

Model ReleasesDGX agent

This post discusses Ollama's functionality for resuming work with Anthropic's Claude models, indicating that users can safely restore previous sessions or states using a command-line interface (`ollam

Language Models Struggle to Use Representations Learned In-Context

ApplicationsDGX agent

arXiv:2602.04212v2 Announce Type: replace Abstract: Though large language models (LLMs) have enabled great success across a wide variety of tasks, they still appear to fall short of one of the loftier

Latent Generative Modeling of Random Fields from Limited Training Data

TutorialsDGX agent

arXiv:2505.13007v2 Announce Type: replace Abstract: The ability to accurately model random fields plays a critical role in science and engineering for problems involving uncertain, spatially-varying q

NorBERTo: A ModernBERT Model Trained for Portuguese with 331 Billion Tokens Corpus

Model ReleasesDGX agent

arXiv:2605.00086v1 Announce Type: new Abstract: High-quality corpora are essential for advancing Natural Language Processing (NLP) in Portuguese. Building on previous encoder-only models such as BERTi

Uncertainty Modeling for Multi-Objective RTA Interception with Distillation Acceleration

ResearchDGX agent

arXiv:2511.05582v2 Announce Type: replace Abstract: Real-Time Auction (RTA) Interception aims to filter out invalid or irrelevant traffic to enhance the integrity and reliability of downstream data. H

What Physics do Data-Driven MoCap-to-Radar Models Learn?

SafetyDGX agent

arXiv:2605.00018v1 Announce Type: new Abstract: Data-driven MoCap-to-radar models generate plausible micro-Doppler spectrograms, but do they actually learn the underlying physics? We introduce a physi

2 May 2026

small milestone: uninstalled the chatgpt app. codex is strict superset now! found something cool - among frontier models, @xai @grok 4.30 is…

Model ReleasesDGX agent

small milestone: uninstalled the chatgpt app. codex is strict superset now! found something cool - among frontier models, @xai @grok 4.30 is the most intelligence per dollar you can get, beating even

1 May 2026

ChipLingo: A Systematic Training Framework for Large Language Models in EDA

Model ReleasesDGX agent

arXiv:2604.27415v1 Announce Type: new Abstract: With the rapid advancement of semiconductor technology, Electronic Design Automation (EDA) has become an increasingly knowledge-intensive and document-d

EDU-CIRCUIT-HW: Evaluating Multimodal Large Language Models on Real-World University-Level STEM Student Handwritten Solutions

Model ReleasesDGX agent

arXiv:2602.00095v3 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) hold significant promise for revolutionizing traditional education and reducing teachers' workload. H

FMCL: Class-Aware Client Clustering with Foundation Model Representations for Heterogeneous Federated Learning

ResearchDGX agent

arXiv:2604.27510v1 Announce Type: cross Abstract: Federated Learning (FL) enables collaborative model training across distributed clients without sharing raw data, yet its performance deteriorates und

HighFM: Towards a Foundation Model for Learning Representations from High-Frequency Earth Observation Data

Model ReleasesDGX agent

arXiv:2604.04306v2 Announce Type: replace-cross Abstract: The increasing frequency and severity of climate related disasters have intensified the need for real time monitoring, early warning, and info

LaST-R1: Reinforcing Action via Adaptive Physical Latent Reasoning for VLA Models

Model ReleasesDGX agent

arXiv:2604.28192v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have increasingly incorporated reasoning mechanisms for complex robotic manipulation. However, existing approaches

RL is a bit of a double edged sword: in known territory performance increases, but in unknown territory the model tends to hallucinate that …

Model ReleasesDGX agent

RL is a bit of a double edged sword: in known territory performance increases, but in unknown territory the model tends to hallucinate that it is performing a completely different task it was trained

The latest crop of models remains below 1% on ARC-AGI-3 -- for now. Where will the scores be by the end of the year?

Model ReleasesDGX agent

The latest crop of models remains below 1% on ARC-AGI-3 -- for now. Where will the scores be by the end of the year? GPT-5.5 & Opus 4.7 on ARC-AGI-3 - GPT-5.5: 0.43% - Opus 4.7: 0.18% We found 3 failu

The new Grok comes in below the latest Chinese open weights models, Grok 4 was at the frontier when released. (& Artificial Analysis: please…

Model ReleasesDGX agent

The new Grok comes in below the latest Chinese open weights models, Grok 4 was at the frontier when released. (& Artificial Analysis: please stop using GDPval-AA which is not a useful test of anything

30 Apr 2026

Benchmarking Deep Learning and Vision Foundation Models for Atypical vs. Normal Mitosis Classification with Cross-Dataset Evaluation

Model ReleasesDGX agent

arXiv:2506.21444v4 Announce Type: replace Abstract: Atypical mitosis marks a deviation in the cell division process that has been shown be an independent prognostic marker for tumor malignancy. Howeve

Benchmarking the Safety of Large Language Models for Robotic Health Attendant Control

SafetyDGX agent

arXiv:2604.26577v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly considered for deployment as the control component of robotic health attendants, yet their safety in this

Beyond the Leaderboard: Rethinking Medical Benchmarks for Large Language Models

Model ReleasesDGX agent

arXiv:2508.04325v2 Announce Type: replace-cross Abstract: Large language models (LLMs) show significant potential in healthcare, prompting numerous benchmarks to evaluate their capabilities. However,

Claude Code is tuned for Claude. Codex is tuned for OpenAI models. Until now, 𝚍𝚎𝚎𝚙𝚊𝚐𝚎𝚗𝚝𝚜 had fixed harness defaults, which meant i…

Model ReleasesDGX agent

Claude Code is tuned for Claude. Codex is tuned for OpenAI models. Until now, 𝚍𝚎𝚎𝚙𝚊𝚐𝚎𝚗𝚝𝚜 had fixed harness defaults, which meant it couldn't take advantage of the provider-specific optimizations that

Generative models on phase space

TutorialsDGX agent

arXiv:2604.02415v2 Announce Type: replace-cross Abstract: Deep generative models such as diffusion and flow matching are powerful machine learning tools capable of learning and sampling from high-dime

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents

AgentsDGX agent

arXiv:2604.26752v1 Announce Type: new Abstract: We present GLM-5V-Turbo, a step toward native foundation models for multimodal agents. As foundation models are increasingly deployed in real environmen

QERNEL: a Scalable Large Electron Model

Model ReleasesDGX agent

arXiv:2604.26018v1 Announce Type: cross Abstract: We introduce QERNEL, a foundational neural wavefunction that variationally solves families of parameterized many-electron Hamiltonians and captures th

Self-Jailbreaking: Language Models Can Reason Themselves Out of Safety Alignment After Benign Reasoning Training

Model ReleasesDGX agent

arXiv:2510.20956v2 Announce Type: replace-cross Abstract: We discover a novel and surprising phenomenon of unintentional misalignment in reasoning language models (RLMs), which we call self-jailbreaki

When to Retrieve During Reasoning: Adaptive Retrieval for Large Reasoning Models

Model ReleasesDGX agent

arXiv:2604.26649v1 Announce Type: cross Abstract: Large reasoning models such as DeepSeek-R1 and OpenAI o1 generate extended chains of thought spanning thousands of tokens, yet their integration with

29 Apr 2026

As models, contexts, and workloads grow, hidden assumptions in inference infrastructure can surface as output anomalies. Reliability require…

Model ReleasesDGX agent

As models, contexts, and workloads grow, hidden assumptions in inference infrastructure can surface as output anomalies. Reliability requires more than throughput, latency, and availability. It also r

EvoTSC: Evolving Feature Learning Models for Time Series Classification via Genetic Programming

Model ReleasesDGX agent

arXiv:2604.25499v1 Announce Type: new Abstract: Time series classification is an important analytical task across diverse domains. However, its practical application is often hindered by the scarcity

Exploring Reasoning Reward Model for Agents

Model ReleasesDGX agent

arXiv:2601.22154v2 Announce Type: replace-cross Abstract: Agentic Reinforcement Learning (Agentic RL) has achieved notable success in enabling agents to perform complex reasoning and tool use. However

Faithfulness-QA: A Counterfactual Entity Substitution Dataset for Training Context-Faithful RAG Models

Model ReleasesDGX agent

arXiv:2604.25313v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) models frequently produce answers grounded in parametric memory rather than the retrieved context, undermining the

From Local to Global: Revisiting Structured Pruning Paradigms for Large Language Models

Model ReleasesDGX agent

arXiv:2510.18030v2 Announce Type: replace Abstract: Structured pruning is a practical approach to deploying large language models (LLMs) efficiently, as it yields compact, hardware-friendly architectu

From Soliloquy to Agora: Memory-Enhanced LLM Agents with Decentralized Debate for Optimization Modeling

AgentsDGX agent

arXiv:2604.25847v1 Announce Type: cross Abstract: Optimization modeling underpins real-world decision-making in logistics, manufacturing, energy, and public services, but reliably solving such problem

MotionBricks: Scalable Real-Time Motions with Modular Latent Generative Model and Smart Primitives

ApplicationsDGX agent

arXiv:2604.24833v1 Announce Type: cross Abstract: Despite transformative advances in generative motion synthesis, real-time interactive motion control remains dominated by traditional techniques. In t

Phase-Associative Memory: Sequence Modeling in Complex Hilbert Space

Model ReleasesDGX agent

arXiv:2604.05030v2 Announce Type: replace Abstract: Experiments probing natural language processing by both humans and LLMs suggest that the meaning of a semantic expression is indeterminate prior to

Three Models of RLHF Annotation: Extension, Evidence, and Authority

SafetyDGX agent

arXiv:2604.25895v1 Announce Type: cross Abstract: Preference-based alignment methods, most prominently Reinforcement Learning with Human Feedback (RLHF), use the judgments of human annotators to shape

28 Apr 2026

A Limit Theory of Foundation Models: A Mathematical Approach to Understanding Emergent Intelligence and Scaling Laws

Model ReleasesDGX agent

arXiv:2604.24037v1 Announce Type: new Abstract: Emergent intelligence have played a major role in the modern AI development. While existing studies primarily rely on empirical observations to characte

A Survey on Split Learning for LLM Fine-Tuning: Models, Systems, and Privacy Optimizations

ResearchDGX agent

arXiv:2604.24468v1 Announce Type: cross Abstract: Fine-tuning unlocks large language models (LLMs) for specialized applications, but its high computational cost often puts it out of reach for resource

Accelerating Frequency Domain Diffusion Models with Error-Feedback Event-Driven Caching

ResearchDGX agent

arXiv:2604.22901v1 Announce Type: new Abstract: Diffusion models achieve remarkable success in time series generation. However, slow inference limits their practical deployment. We propose E^2-CRF (Er

AIPsy-Affect: A Keyword-Free Clinical Stimulus Battery for Mechanistic Interpretability of Emotion in Language Models

Model ReleasesDGX agent

arXiv:2604.23719v1 Announce Type: cross Abstract: Mechanistic interpretability research on emotion in large language models -- linear probing, activation patching, sparse autoencoder (SAE) feature ana

An Information-Geometric Framework for Stability Analysis of Large Language Models under Entropic Stress

SafetyDGX agent

arXiv:2604.24076v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly deployed in high-stakes and operational settings, evaluation strategies based solely on aggregate accur

AutoPyVerifier: Learning Compact Executable Verifiers for Large Language Model Outputs

Model ReleasesDGX agent

arXiv:2604.22937v1 Announce Type: new Abstract: Verification is becoming central to both reinforcement-learning-based training and inference-time control of large language models (LLMs). Yet current v

Benchmarking and Mitigating Sycophancy in Medical Vision Language Models

Model ReleasesDGX agent

arXiv:2509.21979v4 Announce Type: replace-cross Abstract: Visual language models (VLMs) have the potential to transform medical workflows. However, the deployment is limited by sycophancy. Despite thi

Can Multimodal Large Language Models Truly Understand Small Objects?

Model ReleasesDGX agent

arXiv:2604.22884v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have shown promising potential in diverse understanding tasks, e.g., image and video analysis, math and physi

Characterizing Vision-Language-Action Models across XPUs: Constraints and Acceleration for On-Robot Deployment

ResearchDGX agent

arXiv:2604.24447v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are promising for generalist robot control, but on-robot deployment is bottlenecked by real-time inference under t

Contextual Linear Activation Steering of Language Models

ResearchDGX agent

arXiv:2604.24693v1 Announce Type: new Abstract: Linear activation steering is a powerful approach for eliciting the capabilities of large language models and specializing their behavior using limited

Discovering Failure Modes in Vision-Language Models using RL

SafetyDGX agent

arXiv:2604.04733v2 Announce Type: replace-cross Abstract: Vision-language Models (VLMs), despite achieving strong performance on multimodal benchmarks, often misinterpret straightforward visual concep

DO-Bench: An Attributable Benchmark for Diagnosing Object Hallucination in Vision-Language Models

Model ReleasesDGX agent

arXiv:2604.22822v1 Announce Type: cross Abstract: Object level hallucination remains a central reliability challenge for vision language models (VLMs), particularly in binary object existence verifica

EAGLE: Expert-Augmented Attention Guidance for Tuning-Free Industrial Anomaly Detection in Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2602.17419v3 Announce Type: replace Abstract: Multimodal large language models (MLLMs) can enrich industrial anomaly detection with semantic descriptions and anomaly reasoning, but they still la

FedRef: Bayesian Fine-Tuning using a Reference Model to Mitigate Catastrophic Forgetting for Heterogeneous Federated Learning

ApplicationsDGX agent

arXiv:2506.23210v5 Announce Type: replace-cross Abstract: Federated learning (FL) enables collaborative model training across distributed clients while preserving data privacy. However, data and syste

Fine-tuning vs. In-context Learning in Large Language Models: A Formal Language Learning Perspective

TutorialsDGX agent

arXiv:2604.23267v1 Announce Type: new Abstract: Large language models (LLMs) operate in two fundamental learning modes - fine-tuning (FT) and in-context learning (ICL) - raising key questions about wh

HeadRouter: Dynamic Head-Weight Routing for Task-Adaptive Audio Token Pruning in Large Audio Language Models

ResearchDGX agent

arXiv:2604.23717v1 Announce Type: cross Abstract: Recent large audio language models (LALMs) demonstrate remarkable capabilities in processing extended multi-modal sequences, yet incur high inference

HeiSD: Hybrid Speculative Decoding for Embodied Vision-Language-Action Models with Kinematic Awareness

ApplicationsDGX agent

arXiv:2603.17573v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) Models have become the mainstream solution for robot control, but suffer from slow inference speeds. Speculative

Instruction-Free Tuning of Large Vision Language Models for Medical Instruction Following

ResearchDGX agent

arXiv:2603.19482v2 Announce Type: replace Abstract: Large vision language models (LVLMs) have demonstrated impressive performance across a wide range of tasks. These capabilities largely stem from vis

Inverting Foundation Models of Brain Function with Simulation-Based Inference

TutorialsDGX agent

arXiv:2604.23865v1 Announce Type: cross Abstract: Foundation models of brain activity promise a new frontier for in silico neuroscience by emulating neural responses to complex stimuli across tasks an

Large Language Models as Virtual Survey Respondents: Evaluating Sociodemographic Response Generation

Model ReleasesDGX agent

arXiv:2509.06337v2 Announce Type: replace Abstract: Questionnaire-based surveys are foundational to social science research and public policymaking, yet traditional survey methods remain costly, time-

Less Is More: Engineering Challenges of On-Device Small Language Model Integration in a Mobile Application

Model ReleasesDGX agent

arXiv:2604.24636v1 Announce Type: cross Abstract: On-device Small Language Models (SLMs) promise fully offline, private AI experiences for mobile users (no cloud dependency, no data leaving the device

Live Knowledge Tracing: Real-Time Adaptation using Tabular Foundation Models

ResearchDGX agent

arXiv:2602.06542v3 Announce Type: replace Abstract: Deep knowledge tracing models have achieved significant breakthroughs in modeling student learning trajectories. However, these architectures requir

← Previous
1…9192939495…1009
Next →