AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,521 results
Safety

Gradient-Gated DPO: Stabilizing Preference Optimization in Language Models

DGX agent

arXiv:2605.02626v1 Announce Type: new Abstract: Preference optimization has become a central paradigm for aligning large language models with human feedback. Direct Preference Optimization (DPO) simpl

safetyarxiv-cs-lg
5 May 2026
Agents
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

llms are getting expensive why we need oss models

DGX agent

Large language models are becoming increasingly costly to develop and operate, creating a need for open-source alternatives that reduce dependency on expensive proprietary models and make AI more acce

agentsharrison-chase--x
5 May 2026
Model Releases

OralMLLM-Bench: Evaluating Cognitive Capabilities of Multimodal Large Language Models in Dental Practice

DGX agent

arXiv:2605.01333v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have emerged as a promising paradigm for dental image analysis. However, their ability to capture the multi-lev

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Seeing the Scene Matters: Revealing Forgetting in Video Understanding Models with a Scene-Aware Long-Video Benchmark

DGX agent

arXiv:2603.27259v2 Announce Type: replace Abstract: Long video understanding (LVU) remains a core challenge in multimodal learning. Although recent vision-language models (VLMs) have made notable prog

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

SteeringDiffusion: A Bottlenecked Activation Control Interface for Diffusion Models

DGX agent

arXiv:2605.01653v1 Announce Type: new Abstract: We introduce SteeringDiffusion, a bottlenecked activation-level control interface for diffusion models that exposes a smooth, monotonic, and runtime-adj

model-releasesarxiv-cs-cv
5 May 2026
Safety

TRAP: Tail-aware Ranking Attack for World-Model Planning

DGX agent

arXiv:2605.01950v1 Announce Type: new Abstract: World models enable long-horizon planning by internally generating and evaluating imagined trajectories, making them a promising foundation for generali

safetyarxiv-cs-lg
5 May 2026
Hardware

ViM-Q: Scalable Algorithm-Hardware Co-Design for Vision Mamba Model Inference on FPGA

DGX agent

arXiv:2605.01935v1 Announce Type: cross Abstract: Vision Mamba (ViM) models offer a compelling efficiency advantage over Transformers by leveraging the linear complexity of State Space Models (SSMs),

hardwarearxiv-cs-cv
5 May 2026
Agents

what model are you choosing for coding tasks?

DGX agent

This post likely discusses Harrison Chase's preferred language model or AI system for handling coding tasks, potentially comparing different models' capabilities for programming work. As the creator o

agentsharrison-chase--x
5 May 2026
Local Ai

AirFM-DDA: Air-Interface Foundation Model in the Delay-Doppler-Angle Domain for AI-Native 6G

DGX agent

arXiv:2605.00020v1 Announce Type: new Abstract: The success of large foundation models is catalyzing a new paradigm for AI-native 6G network design: wireless foundation models for physical layer desig

local-aiarxiv-cs-lg
4 May 2026
Model Releases

AlphaInventory: Evolving White-Box Inventory Policies via Large Language Models with Deployment Guarantees

DGX agent

arXiv:2605.00369v1 Announce Type: new Abstract: We study how large language models can be used to evolve inventory policies in online, non-stationary environments. Our work is motivated by recent adva

model-releasesarxiv-cs-lg
4 May 2026
Agents

An End-to-End Decision-Aware Multi-Scale Attention-Based Model for Explainable Autonomous Driving

DGX agent

arXiv:2605.00291v1 Announce Type: new Abstract: The application of computer vision is gradually increasing across various domains. They employ deep learning models with a black-box nature. Without the

agentsarxiv-cs-cv
4 May 2026
Model Releases

At any point in time, you can safely resume to using Anthropic's models: ollama launch claude-desktop --restore

DGX agent

This post discusses Ollama's functionality for resuming work with Anthropic's Claude models, indicating that users can safely restore previous sessions or states using a command-line interface (`ollam

model-releasesollama--x
4 May 2026
Applications

Language Models Struggle to Use Representations Learned In-Context

DGX agent

arXiv:2602.04212v2 Announce Type: replace Abstract: Though large language models (LLMs) have enabled great success across a wide variety of tasks, they still appear to fall short of one of the loftier

applicationsarxiv-cs-cl
4 May 2026
Tutorials

Latent Generative Modeling of Random Fields from Limited Training Data

DGX agent

arXiv:2505.13007v2 Announce Type: replace Abstract: The ability to accurately model random fields plays a critical role in science and engineering for problems involving uncertain, spatially-varying q

tutorialsarxiv-cs-lg
4 May 2026
Model Releases

NorBERTo: A ModernBERT Model Trained for Portuguese with 331 Billion Tokens Corpus

DGX agent

arXiv:2605.00086v1 Announce Type: new Abstract: High-quality corpora are essential for advancing Natural Language Processing (NLP) in Portuguese. Building on previous encoder-only models such as BERTi

model-releasesarxiv-cs-cl
4 May 2026
Research

Uncertainty Modeling for Multi-Objective RTA Interception with Distillation Acceleration

DGX agent

arXiv:2511.05582v2 Announce Type: replace Abstract: Real-Time Auction (RTA) Interception aims to filter out invalid or irrelevant traffic to enhance the integrity and reliability of downstream data. H

researcharxiv-cs-lg
4 May 2026
Safety

What Physics do Data-Driven MoCap-to-Radar Models Learn?

DGX agent

arXiv:2605.00018v1 Announce Type: new Abstract: Data-driven MoCap-to-radar models generate plausible micro-Doppler spectrograms, but do they actually learn the underlying physics? We introduce a physi

safetyarxiv-cs-lg
4 May 2026
Model Releases

small milestone: uninstalled the chatgpt app. codex is strict superset now! found something cool - among frontier models, @xai @grok 4.30 is…

DGX agent

small milestone: uninstalled the chatgpt app. codex is strict superset now! found something cool - among frontier models, @xai @grok 4.30 is the most intelligence per dollar you can get, beating even

model-releasesswyx--x
2 May 2026
Model Releases

ChipLingo: A Systematic Training Framework for Large Language Models in EDA

DGX agent

arXiv:2604.27415v1 Announce Type: new Abstract: With the rapid advancement of semiconductor technology, Electronic Design Automation (EDA) has become an increasingly knowledge-intensive and document-d

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

EDU-CIRCUIT-HW: Evaluating Multimodal Large Language Models on Real-World University-Level STEM Student Handwritten Solutions

DGX agent

arXiv:2602.00095v3 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) hold significant promise for revolutionizing traditional education and reducing teachers' workload. H

model-releasesarxiv-cs-ai
1 May 2026
Research

FMCL: Class-Aware Client Clustering with Foundation Model Representations for Heterogeneous Federated Learning

DGX agent

arXiv:2604.27510v1 Announce Type: cross Abstract: Federated Learning (FL) enables collaborative model training across distributed clients without sharing raw data, yet its performance deteriorates und

researcharxiv-cs-cv
1 May 2026
Model Releases

HighFM: Towards a Foundation Model for Learning Representations from High-Frequency Earth Observation Data

DGX agent

arXiv:2604.04306v2 Announce Type: replace-cross Abstract: The increasing frequency and severity of climate related disasters have intensified the need for real time monitoring, early warning, and info

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

LaST-R1: Reinforcing Action via Adaptive Physical Latent Reasoning for VLA Models

DGX agent

arXiv:2604.28192v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have increasingly incorporated reasoning mechanisms for complex robotic manipulation. However, existing approaches

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

RL is a bit of a double edged sword: in known territory performance increases, but in unknown territory the model tends to hallucinate that …

DGX agent

RL is a bit of a double edged sword: in known territory performance increases, but in unknown territory the model tends to hallucinate that it is performing a completely different task it was trained

model-releasesfrancois-chollet--x
1 May 2026
Model Releases

The latest crop of models remains below 1% on ARC-AGI-3 -- for now. Where will the scores be by the end of the year?

DGX agent

The latest crop of models remains below 1% on ARC-AGI-3 -- for now. Where will the scores be by the end of the year? GPT-5.5 & Opus 4.7 on ARC-AGI-3 - GPT-5.5: 0.43% - Opus 4.7: 0.18% We found 3 failu

model-releasesfrancois-chollet--x
1 May 2026
Model Releases

The new Grok comes in below the latest Chinese open weights models, Grok 4 was at the frontier when released. (& Artificial Analysis: please…

DGX agent

The new Grok comes in below the latest Chinese open weights models, Grok 4 was at the frontier when released. (& Artificial Analysis: please stop using GDPval-AA which is not a useful test of anything

model-releasesethan-mollick--x
1 May 2026
Model Releases

Benchmarking Deep Learning and Vision Foundation Models for Atypical vs. Normal Mitosis Classification with Cross-Dataset Evaluation

DGX agent

arXiv:2506.21444v4 Announce Type: replace Abstract: Atypical mitosis marks a deviation in the cell division process that has been shown be an independent prognostic marker for tumor malignancy. Howeve

model-releasesarxiv-cs-cv
30 Apr 2026
Safety

Benchmarking the Safety of Large Language Models for Robotic Health Attendant Control

DGX agent

arXiv:2604.26577v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly considered for deployment as the control component of robotic health attendants, yet their safety in this

safetyarxiv-cs-ai
30 Apr 2026
Model Releases

Beyond the Leaderboard: Rethinking Medical Benchmarks for Large Language Models

DGX agent

arXiv:2508.04325v2 Announce Type: replace-cross Abstract: Large language models (LLMs) show significant potential in healthcare, prompting numerous benchmarks to evaluate their capabilities. However,

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Claude Code is tuned for Claude. Codex is tuned for OpenAI models. Until now, 𝚍𝚎𝚎𝚙𝚊𝚐𝚎𝚗𝚝𝚜 had fixed harness defaults, which meant i…

DGX agent

Claude Code is tuned for Claude. Codex is tuned for OpenAI models. Until now, 𝚍𝚎𝚎𝚙𝚊𝚐𝚎𝚗𝚝𝚜 had fixed harness defaults, which meant it couldn't take advantage of the provider-specific optimizations that

model-releasesharrison-chase--x
30 Apr 2026
Tutorials

Generative models on phase space

DGX agent

arXiv:2604.02415v2 Announce Type: replace-cross Abstract: Deep generative models such as diffusion and flow matching are powerful machine learning tools capable of learning and sampling from high-dime

tutorialsarxiv-cs-ai
30 Apr 2026
Agents

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents

DGX agent

arXiv:2604.26752v1 Announce Type: new Abstract: We present GLM-5V-Turbo, a step toward native foundation models for multimodal agents. As foundation models are increasingly deployed in real environmen

agentsarxiv-cs-cv
30 Apr 2026
Model Releases

QERNEL: a Scalable Large Electron Model

DGX agent

arXiv:2604.26018v1 Announce Type: cross Abstract: We introduce QERNEL, a foundational neural wavefunction that variationally solves families of parameterized many-electron Hamiltonians and captures th

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Self-Jailbreaking: Language Models Can Reason Themselves Out of Safety Alignment After Benign Reasoning Training

DGX agent

arXiv:2510.20956v2 Announce Type: replace-cross Abstract: We discover a novel and surprising phenomenon of unintentional misalignment in reasoning language models (RLMs), which we call self-jailbreaki

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

When to Retrieve During Reasoning: Adaptive Retrieval for Large Reasoning Models

DGX agent

arXiv:2604.26649v1 Announce Type: cross Abstract: Large reasoning models such as DeepSeek-R1 and OpenAI o1 generate extended chains of thought spanning thousands of tokens, yet their integration with

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

As models, contexts, and workloads grow, hidden assumptions in inference infrastructure can surface as output anomalies. Reliability require…

DGX agent

As models, contexts, and workloads grow, hidden assumptions in inference infrastructure can surface as output anomalies. Reliability requires more than throughput, latency, and availability. It also r

model-releaseszhipu-ai--x
29 Apr 2026
Model Releases

EvoTSC: Evolving Feature Learning Models for Time Series Classification via Genetic Programming

DGX agent

arXiv:2604.25499v1 Announce Type: new Abstract: Time series classification is an important analytical task across diverse domains. However, its practical application is often hindered by the scarcity

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

Exploring Reasoning Reward Model for Agents

DGX agent

arXiv:2601.22154v2 Announce Type: replace-cross Abstract: Agentic Reinforcement Learning (Agentic RL) has achieved notable success in enabling agents to perform complex reasoning and tool use. However

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Faithfulness-QA: A Counterfactual Entity Substitution Dataset for Training Context-Faithful RAG Models

DGX agent

arXiv:2604.25313v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) models frequently produce answers grounded in parametric memory rather than the retrieved context, undermining the

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

From Local to Global: Revisiting Structured Pruning Paradigms for Large Language Models

DGX agent

arXiv:2510.18030v2 Announce Type: replace Abstract: Structured pruning is a practical approach to deploying large language models (LLMs) efficiently, as it yields compact, hardware-friendly architectu

model-releasesarxiv-cs-cl
29 Apr 2026
Agents

From Soliloquy to Agora: Memory-Enhanced LLM Agents with Decentralized Debate for Optimization Modeling

DGX agent

arXiv:2604.25847v1 Announce Type: cross Abstract: Optimization modeling underpins real-world decision-making in logistics, manufacturing, energy, and public services, but reliably solving such problem

agentsarxiv-cs-lg
29 Apr 2026
Applications

MotionBricks: Scalable Real-Time Motions with Modular Latent Generative Model and Smart Primitives

DGX agent

arXiv:2604.24833v1 Announce Type: cross Abstract: Despite transformative advances in generative motion synthesis, real-time interactive motion control remains dominated by traditional techniques. In t

applicationsarxiv-cs-lg
29 Apr 2026
Model Releases

Phase-Associative Memory: Sequence Modeling in Complex Hilbert Space

DGX agent

arXiv:2604.05030v2 Announce Type: replace Abstract: Experiments probing natural language processing by both humans and LLMs suggest that the meaning of a semantic expression is indeterminate prior to

model-releasesarxiv-cs-cl
29 Apr 2026
Safety

Three Models of RLHF Annotation: Extension, Evidence, and Authority

DGX agent

arXiv:2604.25895v1 Announce Type: cross Abstract: Preference-based alignment methods, most prominently Reinforcement Learning with Human Feedback (RLHF), use the judgments of human annotators to shape

safetyarxiv-cs-cl
29 Apr 2026
Model Releases

A Limit Theory of Foundation Models: A Mathematical Approach to Understanding Emergent Intelligence and Scaling Laws

DGX agent

arXiv:2604.24037v1 Announce Type: new Abstract: Emergent intelligence have played a major role in the modern AI development. While existing studies primarily rely on empirical observations to characte

model-releasesarxiv-cs-lg
28 Apr 2026
Research

A Survey on Split Learning for LLM Fine-Tuning: Models, Systems, and Privacy Optimizations

DGX agent

arXiv:2604.24468v1 Announce Type: cross Abstract: Fine-tuning unlocks large language models (LLMs) for specialized applications, but its high computational cost often puts it out of reach for resource

researcharxiv-cs-cl
28 Apr 2026
Research

Accelerating Frequency Domain Diffusion Models with Error-Feedback Event-Driven Caching

DGX agent

arXiv:2604.22901v1 Announce Type: new Abstract: Diffusion models achieve remarkable success in time series generation. However, slow inference limits their practical deployment. We propose E^2-CRF (Er

researcharxiv-cs-lg
28 Apr 2026
Model Releases

AIPsy-Affect: A Keyword-Free Clinical Stimulus Battery for Mechanistic Interpretability of Emotion in Language Models

DGX agent

arXiv:2604.23719v1 Announce Type: cross Abstract: Mechanistic interpretability research on emotion in large language models -- linear probing, activation patching, sparse autoencoder (SAE) feature ana

model-releasesarxiv-cs-ai
28 Apr 2026
← Previous
1…114115116117118…1261
Next →