AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
AllBlog
90,316Total entries
1Added by human
90,315Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,172 results
Model Releases

Visual Latents Know More Than They Say: Unsilencing Latent Reasoning in MLLMs

DGX agent

arXiv:2605.02735v1 Announce Type: new Abstract: Continuous latent-space reasoning offers a compact alternative to textual chain-of-thought for multimodal models, enabling high-dimensional visual evide

model-releasesarxiv-cs-lg
5 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

AGoQ: Activation and Gradient Quantization for Memory-Efficient Distributed Training of LLMs

DGX agent

arXiv:2605.00539v1 Announce Type: new Abstract: Quantization is a key method for reducing the GPU memory requirement of training large language models (LLMs). Yet, current approaches are ineffective f

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

BanglaSocialBench: A Benchmark for Evaluating Sociopragmatic and Cultural Alignment of LLMs in Bangladeshi Social Interaction

DGX agent

arXiv:2603.15949v3 Announce Type: replace Abstract: Large Language Models have demonstrated strong multilingual fluency, yet fluency alone does not guarantee socially appropriate language use. In high

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Bring Your Own Prompts: Use-Case-Specific Bias and Fairness Evaluation for LLMs

DGX agent

arXiv:2407.10853v5 Announce Type: replace Abstract: Bias and fairness risks in Large Language Models (LLMs) vary substantially across deployment contexts, yet existing approaches lack systematic guida

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Can Coding Agents Reproduce Findings in Computational Materials Science?

DGX agent

arXiv:2605.00803v1 Announce Type: cross Abstract: Large language models are increasingly deployed as autonomous coding agents and have achieved remarkably strong performance on software engineering be

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

ControBench: An Interaction-Aware Benchmark for Controversial Discourse Analysis on Social Networks

DGX agent

arXiv:2605.00513v1 Announce Type: new Abstract: Understanding how people argue across ideological divides online is important for studying political polarization, misinformation, and content moderatio

model-releasesarxiv-cs-cl
4 May 2026
Safety

Disentangled Safety Adapters Enable Efficient Guardrails and Flexible Inference-Time Alignment

DGX agent

arXiv:2506.00166v2 Announce Type: replace-cross Abstract: Existing paradigms for ensuring AI safety, such as guardrail models and alignment training, often compromise either inference efficiency or de

safetyarxiv-cs-cl
4 May 2026
Model Releases

Documentation https://docs.ollama.com/integrations/claude-desktop

DGX agent

This documentation page describes how to integrate Ollama with Claude Desktop, enabling users to run local language models through the Anthropic Claude interface. The integration allows Claude Desktop

model-releasesollama--x
4 May 2026
Model Releases

Granite 4.1 3B SVG Pelican Gallery

DGX agent

Granite 4.1 3B SVG Pelican Gallery IBM released their Granite 4.1 family of LLMs a few days ago. They're Apache 2.0 licensed and come in 3B, 8B and 30B sizes. Granite 4.1 LLMs: How They’re Built by Gr

model-releasessimon-willison
4 May 2026
Model Releases

How Frontier LLMs Adapt to Neurodivergence Context: A Measurement Framework for Surface vs. Structural Change in System-Prompted Responses

DGX agent

arXiv:2605.00113v1 Announce Type: new Abstract: We examine if frontier chat-based large language models (LLMs) adjust their outputs based on neurodivergence (ND) context in system prompts and describe

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

LWiAI Podcast #243 - GPT 5.5, DeepSeek V4, AI safety sabotage

DGX agent

This podcast episode from Last Week in AI discusses recent developments in large language models, including updates on GPT 5.5 and DeepSeek V4, while also covering concerns about potential sabotage or

model-releaseslast-week-in-ai
4 May 2026
Model Releases

Make Your LVLM KV Cache More Lightweight

DGX agent

arXiv:2605.00789v1 Announce Type: new Abstract: Key-Value (KV) cache has become a de facto component of modern Large Vision-Language Models (LVLMs) for inference. While it enhances decoding efficiency

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Minimizing Human Intervention in Online Classification

DGX agent

arXiv:2510.23557v2 Announce Type: replace-cross Abstract: Training or fine-tuning large language model (LLM)-based systems often requires costly human feedback, yet there is limited understanding of h

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

PEACE: Cross-modal Enhanced Pediatric-Adult ECG Alignment for Robust Pediatric Diagnosis

DGX agent

arXiv:2605.00647v1 Announce Type: new Abstract: Automated pediatric electrocardiogram (ECG) diagnosis remains challenging because models trained predominantly on adult data suffer from substantial cro

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs

DGX agent

arXiv:2605.00814v1 Announce Type: new Abstract: While autoregressive Large Vision-Language Models (LVLMs) demonstrate remarkable proficiency in multimodal tasks, they face a 'Visual Signal Dilution' p

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Reasoning-Intensive Regression

DGX agent

arXiv:2508.21762v3 Announce Type: replace Abstract: AI researchers and practitioners increasingly apply large language models (LLMs) to what we call reasoning-intensive regression (RiR), i.e., deducin

model-releasesarxiv-cs-cl
4 May 2026
Research

RouteProfile: Elucidating the Design Space of LLM Profiles for Routing

DGX agent

arXiv:2605.00180v1 Announce Type: cross Abstract: As the large language model (LLM) ecosystem expands, individual models exhibit varying capabilities across queries, benchmarks, and domains, motivatin

researcharxiv-cs-cl
4 May 2026
Model Releases

Smart Profit-Aware Crop Advisory System: Kisan AI

DGX agent

arXiv:2605.00133v1 Announce Type: new Abstract: Modern crop advisory systems exhibit a critical limitation termed extit{economic blindness}. These systems primarily optimize for biological yield, ofte

model-releasesarxiv-cs-lg
4 May 2026
Safety

Statistical Impossibility and Possibility of Aligning LLMs with Human Preferences: From Condorcet Paradox to Nash Equilibrium

DGX agent

arXiv:2503.10990v2 Announce Type: replace-cross Abstract: Aligning large language models (LLMs) with diverse human preferences is critical for ensuring fairness and informed outcomes when deploying th

safetyarxiv-cs-lg
4 May 2026
Applications

The Power of Order: Fooling LLMs with Adversarial Table Permutations

DGX agent

arXiv:2605.00445v1 Announce Type: new Abstract: Large Language Models have achieved remarkable success and are increasingly deployed in critical applications involving tabular data, such as Table Ques

applicationsarxiv-cs-lg
4 May 2026
Model Releases

Why Do LLMs Struggle in Strategic Play? Broken Links Between Observations, Beliefs, and Actions

DGX agent

arXiv:2605.00226v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly tasked with strategic decision-making under incomplete information, such as in negotiation and policymakin

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

I am not sure I would agree with all of this, but the relationship between Anthropic and Claude is quite different than the relationship bet…

DGX agent

I am not sure I would agree with all of this, but the relationship between Anthropic and Claude is quite different than the relationship between other labs and their models. And that shows up in lots

model-releasesethan-mollick--x
3 May 2026
Model Releases

Its getting hard to benchmark frontier agent performance on longer tasks. Repeated measurement is very expensive and there are differences b…

DGX agent

Its getting hard to benchmark frontier agent performance on longer tasks. Repeated measurement is very expensive and there are differences between using models in harnesses versus via APIs. I suspect

model-releasesethan-mollick--x
3 May 2026
Model Releases

Public interpretability dataset and benchmark library for a novel transformer architecture [R]

DGX agent

This work presents an explainability library for transformer models that provides tools for understanding transformer behavior through attributions and concept-based explanations . The resource likely

model-releasesr-machinelearning
3 May 2026
Model Releases

SULPHUR 2 RELEASED

DGX agent

Stable Diffusion 2.0 is an open-source text-to-image model that includes improved text-to-image capabilities using the OpenCLIP encoder, generating higher quality images at resolutions of 512x512 and

model-releasesr-stablediffusion
3 May 2026
Model Releases

(Sorry, after seeing so many of these, could not resist): 🚨 BREAKING: Google just dropped a NEW paper that completely deletes RNNs from exi…

DGX agent

(Sorry, after seeing so many of these, could not resist): 🚨 BREAKING: Google just dropped a NEW paper that completely deletes RNNs from existence. No recurrence. No convolutions. Nothing. Just one mec

model-releasesethan-mollick--x
2 May 2026
Model Releases

BatteryPass-12K: The First Dataset for the Novel Digital Battery Passport Conformance Task

DGX agent

arXiv:2604.26986v1 Announce Type: new Abstract: We introduce a novel task of digital battery passport (DBP) conformance classification and introduce the first public benchmark for the task: BatteryPas

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

Claw-Eval-Live: A Live Agent Benchmark for Evolving Real-World Workflows

DGX agent

arXiv:2604.28139v1 Announce Type: cross Abstract: LLM agents are expected to complete end-to-end units of work across software tools, business services, and local workspaces. Yet many agent benchmarks

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

DEFault++: Automated Fault Detection, Categorization, and Diagnosis for Transformer Architectures

DGX agent

arXiv:2604.28118v1 Announce Type: cross Abstract: Transformer models are widely deployed in critical AI applications, yet faults in their attention mechanisms, projections, and other internal componen

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Exploring Interaction Paradigms for LLM Agents in Scientific Visualization

DGX agent

arXiv:2604.27996v1 Announce Type: new Abstract: This paper examines how different types of large language model (LLM) agents perform on scientific visualization (SciVis) tasks, where users generate vi

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

From Test-taking to Cognitive Scaffolding: A Pedagogical Diagnostic Benchmark for LLMs on English Standardized Tests

DGX agent

arXiv:2505.17056v2 Announce Type: replace-cross Abstract: As large language models (LLMs) are increasingly integrated into educational tools, current evaluations on standardized tests predominantly fo

model-releasesarxiv-cs-ai
1 May 2026
Research

Generative Human Geometry Distribution

DGX agent

arXiv:2503.01448v5 Announce Type: replace Abstract: Realistic human geometry generation is an important yet challenging task, requiring both the preservation of fine clothing details and the accurate

researcharxiv-cs-cv
1 May 2026
Model Releases

GlowQ: Group-Shared LOw-Rank Approximation for Quantized LLMs

DGX agent

arXiv:2603.25385v2 Announce Type: replace-cross Abstract: Quantization techniques such as BitsAndBytes, AWQ, and GPTQ are widely used as a standard method in deploying large language models but often

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Intent2Tx: Benchmarking LLMs for Translating Natural Language Intents into Ethereum Transactions

DGX agent

arXiv:2604.27763v1 Announce Type: new Abstract: The emergence of Large Language Models (LLMs) offers a transformative interface for Web3, yet existing benchmarks fail to capture the complexity of tran

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

InteractWeb-Bench: Can Multimodal Agent Escape Blind Execution in Interactive Website Generation?

DGX agent

arXiv:2604.27419v1 Announce Type: new Abstract: With the advancement of multimodal large language models (MLLMs) and coding agents, the website development has shifted from manual programming to agent

model-releasesarxiv-cs-ai
1 May 2026
Applications

Making Logic a First-Class Citizen in Generative ML for Networking

DGX agent

arXiv:2506.23964v3 Announce Type: replace-cross Abstract: Generative ML models are increasingly popular in networking for tasks such as telemetry imputation, prediction, and synthetic trace generation

applicationsarxiv-cs-lg
1 May 2026
Model Releases

MCPHunt: An Evaluation Framework for Cross-Boundary Data Propagation in Multi-Server MCP Agents

DGX agent

arXiv:2604.27819v1 Announce Type: new Abstract: Multi-server MCP agents create an information-flow control problem: faithful tool composition can turn individually benign read/write permissions into c

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

ORFS-agent: Tool-Using Agents for Chip Design Optimization

DGX agent

arXiv:2506.08332v3 Announce Type: replace Abstract: Machine learning has been widely used to optimize complex engineering workflows across numerous domains. In integrated circuit design, modern flows

model-releasesarxiv-cs-ai
1 May 2026
Applications

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation

DGX agent

arXiv:2604.27747v1 Announce Type: cross Abstract: Large language model (LLM)-based generative list-wise recommendation has advanced rapidly, but decoding remains sequential and thus latency-prone. To

applicationsarxiv-cs-ai
1 May 2026
Model Releases

Post-Optimization Adaptive Rank Allocation for LoRA

DGX agent

arXiv:2604.27796v1 Announce Type: new Abstract: Exponential growth in the scale of modern foundation models has led to the widespread adoption of Low-Rank Adaptation (LoRA) as a parameter-efficient fi

model-releasesarxiv-cs-ai
1 May 2026
Applications

Predicting Covariate-Driven Spatial Deformation for Nonstationary Gaussian Processes

DGX agent

arXiv:2604.27280v1 Announce Type: new Abstract: Nonstationary Gaussian processes (GPs) are essential for modeling complex, locally heterogeneous spatial data. A common modeling approach is the spatial

applicationsarxiv-cs-lg
1 May 2026
Model Releases

Reinforced Agent: Inference-Time Feedback for Tool-Calling Agents

DGX agent

arXiv:2604.27233v1 Announce Type: new Abstract: Tool-calling agents are evaluated on tool selection, parameter accuracy, and scope recognition, yet LLM trajectory assessments remain inherently post-ho

model-releasesarxiv-cs-ai
1 May 2026
Tutorials

Sequential Inference for Gaussian Processes: A Signal Processing Perspective

DGX agent

arXiv:2604.28163v1 Announce Type: cross Abstract: The proliferation of capable and efficient machine learning (ML) models marks one of the strongest methodological shifts in signal processing (SP) in

tutorialsarxiv-cs-lg
1 May 2026
Model Releases

Step-level Optimization for Efficient Computer-use Agents

DGX agent

arXiv:2604.27151v1 Announce Type: new Abstract: Computer-use agents provide a promising path toward general software automation because they can interact directly with arbitrary graphical user interfa

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

AdaMem: Adaptive User-Centric Memory for Long-Horizon Dialogue Agents

DGX agent

arXiv:2603.16496v2 Announce Type: replace Abstract: Large language model (LLM) agents increasingly rely on external memory to support long-horizon interaction, personalized assistance, and multi-step

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

Associative-State Universal Transformers: Sparse Retrieval Meets Structured Recurrence

DGX agent

arXiv:2604.25930v1 Announce Type: new Abstract: We study whether a structured recurrent state can serve as a compact associative backbone for language modeling while still supporting exact retrieval.

model-releasesarxiv-cs-cl
30 Apr 2026
Research

ChartVerse: Scaling Chart Reasoning via Reliable Programmatic Synthesis from Scratch

DGX agent

arXiv:2601.13606v2 Announce Type: replace Abstract: Chart reasoning is a critical capability for Vision Language Models (VLMs). However, the development of open-source models is severely hindered by t

researcharxiv-cs-cv
30 Apr 2026
Model Releases

ClassEval-Pro: A Cross-Domain Benchmark for Class-Level Code Generation

DGX agent

arXiv:2604.26923v1 Announce Type: cross Abstract: LLMs have achieved strong results on both function-level code synthesis and repository-level code modification, yet a capability that falls between th

model-releasesarxiv-cs-cl
30 Apr 2026
← Previous
1…468469470471472…1358
Next →