AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,499 results
28 Apr 2026

Hybrid JIT-CUDA Graph Optimization for Low-Latency Large Language Model Inference

Model ReleasesDGX agent

arXiv:2604.23467v1 Announce Type: cross Abstract: Large Language Models (LLMs) have achieved strong performance across natural language and multimodal tasks, yet their practical deployment remains con

I have been playing with the new Outlook agent, and it is fine, but really awkward to use, since you have to ask for things in a chatbot win…

Model ReleasesDGX agent

I have been playing with the new Outlook agent, and it is fine, but really awkward to use, since you have to ask for things in a chatbot window, then go to your drafts, etc. And Claude Cowork does the

I pulled 3,182 tweets analyzing @NousResearch Hermes Agent versus Claude Code to understand exactly why developers are choosing one or the o…


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

I pulled 3,182 tweets analyzing @NousResearch Hermes Agent versus Claude Code to understand exactly why developers are choosing one or the other I want to figure out why people use Hermes vs Claude/ C

i think openai called it fear based marketing but mythos didn't get released because it's too dangerous, which was shortly after the whole d…

Model ReleasesDGX agent

i think openai called it fear based marketing but mythos didn't get released because it's too dangerous, which was shortly after the whole dod thing, and i think trump said something like 'it's possib

I used to think the best way to come up with ways to use AI was to think about a painpoint or a problem and see if AI can make it better. I …

Model ReleasesDGX agent

I used to think the best way to come up with ways to use AI was to think about a painpoint or a problem and see if AI can make it better. I no longer think that. Why? Because that put every person int

'If they ever tell my story let them say that I walked with giants'--Troy I am humbled&excited the model we released last week is trending #…

Model ReleasesDGX agent

'If they ever tell my story let them say that I walked with giants'--Troy I am humbled&excited the model we released last week is trending #2 on @huggingface, between giant models such as DeepSeek,Qwe

I'm so confused…

Model ReleasesDGX agent

I'm so confused… We're excited to partner with Google to offer Grounding With Exa inside of Gemini models! Using Exa's agent-first search, Gemini models can now access billions of websites, technical

ImmerIris: A Large-Scale Dataset and Benchmark for Off-Axis and Unconstrained Iris Recognition in Immersive Applications

Model ReleasesDGX agent

arXiv:2510.10113v3 Announce Type: replace Abstract: Recently, iris recognition is regaining prominence in immersive applications such as extended reality as a means of seamless user identification. Th

IMPA-Net: Meteorology-Aware Multi-Scale Attention and Dynamic Loss for Extreme Convective Radar Nowcasting

Model ReleasesDGX agent

arXiv:2604.24224v1 Announce Type: new Abstract: Short-range prediction of convective precipitation from weather radar observations is essential for severe weather warnings. However, deep learning mode

Improving Vision-language Models with Perception-centric Process Reward Models

Model ReleasesDGX agent

arXiv:2604.24583v1 Announce Type: new Abstract: Recent advancements in reinforcement learning with verifiable rewards (RLVR) have significantly improved the complex reasoning ability of vision-languag

In the last four Claude Code CLI releases, we’ve shipped 50+ stability and performance fixes. Faster resume, stable auth, lower memory, fewe…

Model ReleasesDGX agent

Claude Code CLI has released over 50 stability and performance improvements in its last four updates, including faster resume functionality, more stable authentication, reduced memory usage, and fewer

INHerit-SG: Incremental Hierarchical Semantic Scene Graphs with RAG-Style Retrieval

Model ReleasesDGX agent

arXiv:2602.12971v2 Announce Type: replace Abstract: Driven by recent advancements in foundation models, semantic scene graphs have emerged as a promising paradigm for high-level 3D environmental abstr

InquireMobile: Teaching VLM-based Mobile Agent to Request Human Assistance via Reinforcement Fine-Tuning

Model ReleasesDGX agent

arXiv:2508.19679v2 Announce Type: replace Abstract: Recent advances in Vision-Language Models (VLMs) have enabled mobile agents to perceive and interact with real-world mobile environments based on hu

Insert In Style: A Zero-Shot Generative Framework for Harmonious Cross-Domain Object Composition

Model ReleasesDGX agent

arXiv:2511.15197v2 Announce Type: replace Abstract: Reference-based object composition involves integrating foreground reference image with background scene to produce harmonious fused image. This tas

IntrAgent: An LLM Agent for Content-Grounded Information Retrieval through Literature Review

Model ReleasesDGX agent

arXiv:2604.22861v1 Announce Type: cross Abstract: Scientific research relies on accurate information retrieval from literature to support analytical decisions. In this work, we introduce a new task, I

Introducing NVIDIA Nemotron 3 Nano Omni: Long-Context Multimodal Intelligence for Documents, Audio and Video Agents

Model ReleasesDGX agent

NVIDIA's Nemotron 3 Nano Omni is a lightweight multimodal AI model capable of processing documents, audio, and video inputs for building intelligent agents. The model supports long-context understandi

Introducing talkie: a 13B vintage language model from 1930

Model ReleasesDGX agent

Introducing talkie: a 13B vintage language model from 1930 New project from Nick Levine, David Duvenaud, and Alec Radford (of GPT, GPT-2, Whisper fame). talkie-1930-13b-base (53.1 GB) is a '13B langua

I’ve been saying this for over a year now: Frontier models are fantastic, but the real future is frontier-level models (in every way) runnin…

Model ReleasesDGX agent

I’ve been saying this for over a year now: Frontier models are fantastic, but the real future is frontier-level models (in every way) running locally on your own hardware. I think this is 18-24 months

I've been trying to make transformers more agent-friendly: agentic CLI, a skill, doc rewrites, canonical examples. It felt a bit like shooti…

Model ReleasesDGX agent

I've been trying to make transformers more agent-friendly: agentic CLI, a skill, doc rewrites, canonical examples. It felt a bit like shooting in the dark: hard to measure progress, and hard to ensure

I've been working on a side project for the last few weeks... what if you could have a Gemma powered app that would let you have a personal …

Model ReleasesDGX agent

I've been working on a side project for the last few weeks... what if you could have a Gemma powered app that would let you have a personal assistant that could browse the internet with you, do resear

Jailbreaking Frontier Foundation Models Through Intention Deception

Model ReleasesDGX agent

arXiv:2604.24082v1 Announce Type: cross Abstract: Large (vision-)language models exhibit remarkable capability but remain highly susceptible to jailbreaking. Existing safety training approaches aim to

JudgeSense: A Benchmark for Prompt Sensitivity in LLM-as-a-Judge Systems

Model ReleasesDGX agent

arXiv:2604.23478v1 Announce Type: new Abstract: Large language models are increasingly deployed as automated judges for evaluating other models, yet the stability of their verdicts under semantically

Judging the Judges: A Systematic Evaluation of Bias Mitigation Strategies in LLM-as-a-Judge Pipelines

Model ReleasesDGX agent

arXiv:2604.23178v1 Announce Type: new Abstract: LLM-as-a-Judge has become the dominant paradigm for evaluating language model outputs, yet LLM judges exhibit systematic biases that compromise evaluati

K-MetBench: A Multi-Dimensional Benchmark for Fine-Grained Evaluation of Expert Reasoning, Locality, and Multimodality in Meteorology

Model ReleasesDGX agent

arXiv:2604.24645v1 Announce Type: cross Abstract: The development of practical (multimodal) large language model assistants for Korean weather forecasters is hindered by the absence of a multidimensio

KLong: Training LLM Agent for Extremely Long-horizon Tasks

Model ReleasesDGX agent

arXiv:2602.17547v3 Announce Type: replace Abstract: This paper introduces KLong, an open-source LLM agent trained to solve extremely long-horizon tasks. The principle is to first cold-start the model

Large Language Models as Virtual Survey Respondents: Evaluating Sociodemographic Response Generation

Model ReleasesDGX agent

arXiv:2509.06337v2 Announce Type: replace Abstract: Questionnaire-based surveys are foundational to social science research and public policymaking, yet traditional survey methods remain costly, time-

Latency and Cost of Multi-Agent Intelligent Tutoring at Scale

Model ReleasesDGX agent

arXiv:2604.24110v1 Announce Type: cross Abstract: Multi-agent LLM tutoring systems improve response quality through agent specialization, but each student query triggers several concurrent API calls w

Layerwise Convergence Fingerprints for Runtime Misbehavior Detection in Large Language Models

Model ReleasesDGX agent

arXiv:2604.24542v1 Announce Type: cross Abstract: Large language models deployed at runtime can misbehave in ways that clean-data validation cannot anticipate: training-time backdoors lie dormant unti

Learn how to run a local coding agent! Use: - Pi agent - Gemma 4 26B - Serving engine of choice: e.g. LM Studio

Model ReleasesDGX agent

This resource provides instructions for setting up and running a local coding agent using LM Studio's serving engine, featuring the Pi agent framework and Google's Gemma 4 26B language model. It demon

📚 Learn more here https://mistral.ai/news/workflows

Model ReleasesDGX agent

Mistral AI announced new workflow capabilities or features, likely detailing how users can implement multi-step processes or automation using their AI models and services. The announcement was shared

LearnAlign: Data Selection for LLM Reinforcement Learning with Improved Gradient Alignment

Model ReleasesDGX agent

arXiv:2506.11480v4 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a key technique for enhancing LLMs' reasoning abilities, yet its data ineffic

Learning Gradient-based Mixup with Extrapolation toward Flatter Minima for Domain Generalization

Model ReleasesDGX agent

arXiv:2209.14742v2 Announce Type: replace Abstract: To address distribution shifts between training and test data, domain generalization (DG) leverages multiple source domains to learn a model that ge

Learning in Blocks: A Multi Agent Debate Assisted Personalized Adaptive Learning Framework for Language Learning

Model ReleasesDGX agent

arXiv:2604.22770v1 Announce Type: cross Abstract: Most digital language learning curricula rely on discrete-item quizzes that test recall rather than applied conversational proficiency. When progressi

Learning Latent Graph Geometry via Fixed-Point Schrodinger-Type Activation: A Theoretical Study

Model ReleasesDGX agent

arXiv:2507.20088v3 Announce Type: replace Abstract: We study neural architectures in which each hidden layer is defined by the stationary state of a dissipative Schrodinger-type dynamics on a learned

Learning to Conceal Risk: Controllable Multi-turn Red Teaming for LLMs in the Financial Domain

Model ReleasesDGX agent

arXiv:2509.10546v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed in finance, where unsafe behavior can lead to serious regulatory risks. However, most r

Learning Under Low Illumination: A Dataset and Algorithm for Traffic Sign Recognition

Model ReleasesDGX agent

arXiv:2511.17183v2 Announce Type: replace Abstract: Traffic signboards are vital for road safety and intelligent transportation systems, enabling navigation and autonomous driving. Yet, recognizing tr

LEGO: An LLM Skill-Based Front-End Design Generation Platform

Model ReleasesDGX agent

arXiv:2604.23355v1 Announce Type: new Abstract: Existing LLM-based EDA agents are often isolated task-specific systems. This leads to repeated engineering effort and limited reuse of successful design

Less Is More: Engineering Challenges of On-Device Small Language Model Integration in a Mobile Application

Model ReleasesDGX agent

arXiv:2604.24636v1 Announce Type: cross Abstract: On-device Small Language Models (SLMs) promise fully offline, private AI experiences for mobile users (no cloud dependency, no data leaving the device

Let's talk document formatting. Bold. Italics. Superscripts. Strikethroughs. The visual cues humans rely on every time we read a doc, and on…

Model ReleasesDGX agent

Let's talk document formatting. Bold. Italics. Superscripts. Strikethroughs. The visual cues humans rely on every time we read a doc, and ones existing OCR benchmarks completely ignore. 😱'199' struck

Leveraging LLMs for Multi-File DSL Code Generation: An Industrial Case Study

Model ReleasesDGX agent

arXiv:2604.24678v1 Announce Type: cross Abstract: Large language models (LLMs) perform strongly on general-purpose code generation, yet their applicability to enterprise domain-specific languages (DSL

Lightweight and Production-Ready PDF Visual Element Parsing

Model ReleasesDGX agent

arXiv:2604.23276v1 Announce Type: cross Abstract: PDF documents contain critical visual elements such as figures, tables, and forms whose accurate extraction is essential for document understanding an

Linear-Nonlinear Fusion Neural Operator for Partial Differential Equations

Model ReleasesDGX agent

arXiv:2603.24143v2 Announce Type: replace Abstract: Neural operator learning directly constructs the mapping relationship from the equation parameter space to the solution space, enabling efficient di

LLM-Assisted Op-Amp Behavioral-Level Design via Agentic Human-Mimicking Reasoning

Model ReleasesDGX agent

arXiv:2601.21321v2 Announce Type: replace Abstract: This paper proposes White-Op, an operational amplifier (op-amp) behavioral-level parameter design framework assisted by the human-mimicking reasonin

LLM-Guided Agentic Floor Plan Parsing for Accessible Indoor Navigation of Blind and Low-Vision People

Model ReleasesDGX agent

arXiv:2604.23970v1 Announce Type: new Abstract: Indoor navigation remains a critical accessibility challenge for the blind and low-vision (BLV) individuals, as existing solutions rely on costly per-bu

LLM4SCREENLIT: Recommendations on Assessing the Performance of Large Language Models for Screening Literature in Systematic Reviews

Model ReleasesDGX agent

arXiv:2511.12635v2 Announce Type: replace-cross Abstract: Context: Large language models (LLMs) are increasingly used to screen literature for systematic reviews (SRs), but the standard confusion-matr

LOCAL AI MODELS ARE CATCHING UP TO FRONTIER MODELS WAY FASTER THAN ANYONE EXPECTED this guy ran qwen 3.6 27B locally on a base macbook pro M…

Model ReleasesDGX agent

LOCAL AI MODELS ARE CATCHING UP TO FRONTIER MODELS WAY FASTER THAN ANYONE EXPECTED this guy ran qwen 3.6 27B locally on a base macbook pro M4 with 24GB of memory quantized and stripped of safety guard

Long-Context Aware Upcycling: A New Frontier for Hybrid LLM Scaling

Model ReleasesDGX agent

arXiv:2604.24715v1 Announce Type: new Abstract: Hybrid sequence models that combine efficient Transformer components with linear sequence modeling blocks are a promising alternative to pure Transforme

LongFlow: Efficient KV Cache Compression for Reasoning Models

Model ReleasesDGX agent

arXiv:2603.11504v2 Announce Type: replace-cross Abstract: Recent reasoning models such as OpenAI-o1 and DeepSeek-R1 have shown strong performance on complex tasks including mathematical reasoning and

Looking for the Bottleneck in Fine-grained Temporal Relation Classification

Model ReleasesDGX agent

arXiv:2604.24620v1 Announce Type: new Abstract: Temporal relation classification is the task of determining the temporal relation between pairs of temporal entities in a text. Despite recent advanceme

Lost in Decoding? Reproducing and Stress-Testing the Look-Ahead Prior in Generative Retrieval

Model ReleasesDGX agent

arXiv:2604.23396v1 Announce Type: cross Abstract: Generative retrieval (GR) ranks documents by autoregressively generating document identifiers. Because many GR methods rely on trie-constrained beam s

Lost in the Vibrations: Vision Language Models Fail the Dynamic Gauges Test

Model ReleasesDGX agent

arXiv:2604.22829v1 Announce Type: new Abstract: The digital transformation of industrial manufacturing increasingly relies on the ability of autonomous robots to interact with legacy infrastructure, p

Lovable launches its AI coding app on iOS and Android, letting users code via voice or text AI prompts, and allowing them to switch between a PC and mobile (Sarah Perez/TechCrunch)

Model ReleasesDGX agent

Sarah Perez / TechCrunch: Lovable launches its AI coding app on iOS and Android, letting users code via voice or text AI prompts, and allowing them to switch between a PC and mobile — Apple's recent c

Machine Learning and Deep Learning Models for Short Term Electricity Price Forecasting in Australia's National Electricity Market

Model ReleasesDGX agent

arXiv:2604.23908v1 Announce Type: new Abstract: Short term electricity price forecast is essential in competitive power markets, yet electricity price series exhibit high volatility, irregularity, and

Majorization-Guided Test-Time Adaptation for Vision-Language Models under Modality-Specific Shift

Model ReleasesDGX agent

arXiv:2604.24602v1 Announce Type: new Abstract: Vision-language models transfer well in zero-shot settings, but at deployment the visual and textual branches often shift asymmetrically. Under this con

Mapping License Plate Recoverability Under Extreme Viewing Angles for Oppor-tunistic Urban Sensing

Model ReleasesDGX agent

arXiv:2604.23814v1 Announce Type: cross Abstract: Urban environments contain many imaging sensors built for specific purposes, including ATM, body-worn, CCTV, and dashboard cameras. Under the opportun

MarketBench: Evaluating AI Agents as Market Participants

Model ReleasesDGX agent

arXiv:2604.23897v1 Announce Type: new Abstract: Markets are a promising way to coordinate AI agent activity for similar reasons to those used to justify markets more broadly. In order to effectively p

MEASER: Malware embedding attacks on open-source LLMs

Model ReleasesDGX agent

arXiv:2510.10486v2 Announce Type: replace-cross Abstract: Open-source large language models (LLMs) have demonstrated considerable dominance over proprietary LLMs in resolving neural processing tasks,

Mechanistic Steering of LLMs Reveals Layer-wise Feature Vulnerabilities in Adversarial Settings

Model ReleasesDGX agent

arXiv:2604.23130v1 Announce Type: cross Abstract: Large language models (LLMs) can still be jailbroken into producing harmful outputs despite safety alignment. Existing attacks show this vulnerability

MEG-RAG: Quantifying Multi-modal Evidence Grounding for Evidence Selection in RAG

Model ReleasesDGX agent

arXiv:2604.24564v1 Announce Type: new Abstract: Multimodal Retrieval-Augmented Generation (MRAG) addresses key limitations of Multimodal Large Language Models (MLLMs), such as hallucination and outdat

MEMCoder: Multi-dimensional Evolving Memory for Private-Library-Oriented Code Generation

Model ReleasesDGX agent

arXiv:2604.24222v1 Announce Type: cross Abstract: Large Language Models (LLMs) excel at general code generation, but their performance drops sharply in enterprise settings that rely on internal privat

← Previous
1…305306307308309…375
Next →