AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,292 results
20 Apr 2026

Dynamic Tool Dependency Retrieval for Lightweight Function Calling

Model ReleasesDGX agent

arXiv:2512.17052v4 Announce Type: replace Abstract: Function calling agents powered by Large Language Models (LLMs) select external tools to automate complex tasks. On-device agents typically use a re

Earlier this year Yann LeCun left Meta because Mark Zuckerberg wouldn't bet the company on JEPA. Last week his group dropped the first JEPA …

Model ReleasesDGX agent

Earlier this year Yann LeCun left Meta because Mark Zuckerberg wouldn't bet the company on JEPA. Last week his group dropped the first JEPA that actually trains end-to-end from raw pixels. 15 million

ECG-Lens: Benchmarking ML & DL Models on PTB-XL Dataset

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.15822v1 Announce Type: cross Abstract: Automated classification of electrocardiogram (ECG) signals is a useful tool for diagnosing and monitoring cardiovascular diseases. This study compare

Enabling Predictive Maintenance in District Heating Substations: A Labelled Dataset and Fault Detection Evaluation Framework based on Service Data

Model ReleasesDGX agent

arXiv:2511.14791v2 Announce Type: replace-cross Abstract: Early detection of faults in district heating substations is imperative to reduce return temperatures and enhance efficiency. However, progres

EvoTest: Evolutionary Test-Time Learning for Self-Improving Agentic Systems

Model ReleasesDGX agent

arXiv:2510.13220v2 Announce Type: replace Abstract: A fundamental limitation of current AI agents is their inability to learn complex skills on the fly at test time, often behaving like 'clever but cl

Excited to go public with Logos, a first principles system designed for intuition & innovation We are starting by sharing results in physics…

Model ReleasesDGX agent

Excited to go public with Logos, a first principles system designed for intuition & innovation We are starting by sharing results in physics before moving to other domains Today’s is a lot of fun, one

EXCLUSIVE: @TheOnion tells me they've struck a (long-awaited) deal to take over Infowars. New @pablofindsout on how America's finest (fake) …

Model ReleasesDGX agent

EXCLUSIVE: @TheOnion tells me they've struck a (long-awaited) deal to take over Infowars. New @pablofindsout on how America's finest (fake) news source trolled Alex Jones, with the blessing of Sandy H

'Excuse me, may I say something...' CoLabScience, A Proactive AI Assistant for Biomedical Discovery and LLM-Expert Collaborations

Model ReleasesDGX agent

arXiv:2604.15588v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) into scientific workflows presents exciting opportunities to accelerate biomedical discovery. However,

Exploring LLM-based Verilog Code Generation with Data-Efficient Fine-Tuning and Testbench Automation

Model ReleasesDGX agent

arXiv:2604.15388v1 Announce Type: cross Abstract: Recent advances in large language models have improved code generation, but their use in hardware description languages is still limited. Moreover, tr

Exploring the Capability Boundaries of LLMs in Mastering of Chinese Chouxiang Language

Model ReleasesDGX agent

arXiv:2604.15841v1 Announce Type: new Abstract: While large language models (LLMs) have achieved remarkable success in general language tasks, their performance on Chouxiang Language, a representative

Facial-Expression-Aware Prompting for Empathetic LLM Tutoring

Model ReleasesDGX agent

arXiv:2604.15336v1 Announce Type: cross Abstract: Large language models (LLMs) enable increasingly capable tutoring-style conversational agents, yet effective tutoring requires sensitivity to learners

FETAL-GAUGE: A Benchmark for Assessing Vision-Language Models in Fetal Ultrasound

Model ReleasesDGX agent

arXiv:2512.22278v2 Announce Type: replace Abstract: The growing demand for prenatal ultrasound imaging has intensified a global shortage of trained sonographers, creating barriers to essential fetal h

FineCog-Nav: Integrating Fine-grained Cognitive Modules for Zero-shot Multimodal UAV Navigation

Model ReleasesDGX agent

arXiv:2604.16298v1 Announce Type: new Abstract: UAV vision-language navigation (VLN) requires an agent to navigate complex 3D environments from an egocentric perspective while following ambiguous mult

Fragile Thoughts: How Large Language Models Handle Chain-of-Thought Perturbations

Model ReleasesDGX agent

arXiv:2603.03332v3 Announce Type: replace-cross Abstract: Chain-of-Thought (CoT) prompting has emerged as a foundational technique for eliciting reasoning from Large Language Models (LLMs), yet the ro

Free ~20-50% tok/s on a local llama.cpp setup if you already have a draft model sharing vocabulary with your main one. Local stack quietly g…

Model ReleasesDGX agent

Free ~20-50% tok/s on a local llama.cpp setup if you already have a draft model sharing vocabulary with your main one. Local stack quietly got faster this weekend https://x.com/TechIno219886/status/20

Frequency-Aware Flow Matching for High-Quality Image Generation

Model ReleasesDGX agent

arXiv:2604.15521v1 Announce Type: new Abstract: Flow matching models have emerged as a powerful framework for realistic image generation by learning to reverse a corruption process that progressively

From Benchmarking to Reasoning: A Dual-Aspect, Large-Scale Evaluation of LLMs on Vietnamese Legal Text

Model ReleasesDGX agent

arXiv:2604.16270v1 Announce Type: cross Abstract: The complexity of Vietnam's legal texts presents a significant barrier to public access to justice. While Large Language Models offer a promising solu

From Zero to Detail: A Progressive Spectral Decoupling Paradigm for UHD Image Restoration with New Benchmark

Model ReleasesDGX agent

arXiv:2604.15654v1 Announce Type: new Abstract: Ultra-high-definition (UHD) image restoration poses unique challenges due to the high spatial resolution, diverse content, and fine-grained structures p

FS-Researcher: Test-Time Scaling for Long-Horizon Research Tasks with File-System-Based Agents

Model ReleasesDGX agent

arXiv:2602.01566v2 Announce Type: replace Abstract: Deep research is emerging as a representative long-horizon task for large language model (LLM) agents. However, long trajectories in deep research o

GTA-2: Benchmarking General Tool Agents from Atomic Tool-Use to Open-Ended Workflows

Model ReleasesDGX agent

arXiv:2604.15715v1 Announce Type: cross Abstract: The development of general-purpose agents requires a shift from executing simple instructions to completing complex, real-world productivity workflows

HarmfulSkillBench: How Do Harmful Skills Weaponize Your Agents?

Model ReleasesDGX agent

arXiv:2604.15415v1 Announce Type: cross Abstract: Large language models (LLMs) have evolved into autonomous agents that rely on open skill ecosystems (e.g., ClawHub and Skills.Rest), hosting numerous

Hero-Mamba: Mamba-based Dual Domain Learning for Underwater Image Enhancement

Model ReleasesDGX agent

arXiv:2604.16266v1 Announce Type: new Abstract: Underwater images often suffer from severe degradation, such as color distortion, low contrast, and blurred details, due to light absorption and scatter

Heterogeneous Sheaf Neural Networks

Model ReleasesDGX agent

arXiv:2409.08036v2 Announce Type: replace Abstract: Heterogeneous graphs, whose nodes and edges may belong to different types and feature spaces, arise in a wide variety of real-world domains such as

HiPreNets: High-Precision Neural Networks through Progressive Training

Model ReleasesDGX agent

arXiv:2506.15064v3 Announce Type: replace Abstract: Deep neural networks are powerful tools for solving nonlinear problems in science and engineering, but training highly accurate models becomes chall

Histogram-based Parameter-efficient Tuning for Passive and Active Sonar Classification

Model ReleasesDGX agent

arXiv:2504.15214v3 Announce Type: replace Abstract: Parameter-efficient transfer learning (PETL) methods adapt large artificial neural networks to downstream tasks without fine-tuning the entire model

HyCal: A Training-Free Prototype Calibration Method for Cross-Discipline Few-Shot Class-Incremental Learning

Model ReleasesDGX agent

arXiv:2604.15678v1 Announce Type: new Abstract: Pretrained Vision-Language Models (VLMs) like CLIP show promise in continual learning, but existing Few-Shot Class-Incremental Learning (FSCIL) methods

HyperGVL: Benchmarking and Improving Large Vision-Language Models in Hypergraph Understanding and Reasoning

Model ReleasesDGX agent

arXiv:2604.15648v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) consistently require new arenas to guide their expanding boundaries, yet their capabilities with hypergraphs remain

I built an MCP bridge that connects AI coding tools (Kiro, Claude, Cursor) to a local Ollama instance — still in development, feedback welcome

Model ReleasesDGX agent

An MCP bridge project that enables integration between AI coding tools (Kiro, Claude, and Cursor) and local Ollama instances for offline model inference. The bridge facilitates communication between t

I upgraded my Claude token counter tool to compare different models and Opus 4.7 does appear to use 1.46x times the tokens for text and up t…

Model ReleasesDGX agent

I upgraded my Claude token counter tool to compare different models and Opus 4.7 does appear to use 1.46x times the tokens for text and up to 3x the tokens for images - it's priced the same as Opus 4.

IA-CLAHE: Image-Adaptive Clip Limit Estimation for CLAHE

Model ReleasesDGX agent

arXiv:2604.16010v1 Announce Type: new Abstract: This paper proposes image-adaptive contrast limited adaptive histogram equalization (IA-CLAHE). Conventional CLAHE is widely used to boost the performan

Imagine if they add document signing to Claude :o

Model ReleasesDGX agent

Imagine if they add document signing to Claude :o In Cowork, Claude can now build live artifacts: dashboards and trackers connected to your apps and files. Open one any time and it refreshes with curr

In Cowork, Claude can now build live artifacts: dashboards and trackers connected to your apps and files. Open one any time and it refreshes…

Model ReleasesDGX agent

Claude's Cowork feature now enables users to create live artifacts such as dashboards and trackers that connect to external apps and files, automatically refreshing when accessed. This functionality a

InstructTable: Improving Table Structure Recognition Through Instructions

Model ReleasesDGX agent

arXiv:2604.02880v2 Announce Type: replace Abstract: Table structure recognition (TSR) holds widespread practical importance by parsing tabular images into structured representations, yet encounters si

Intelligence per picojoule, with @itsclivetime and @dylan522p (0:00) Intro (1:22) What is codesign? (2:49) Codesign example: Swish vs ReLU (…

Model ReleasesDGX agent

Intelligence per picojoule, with @itsclivetime and @dylan522p (0:00) Intro (1:22) What is codesign? (2:49) Codesign example: Swish vs ReLU (4:22) Are DeepSeek papers codesign? (6:45) Predicting where

Intelligent Healthcare Imaging Platform: A VLM-Based Framework for Automated Medical Image Analysis and Clinical Report Generation

Model ReleasesDGX agent

arXiv:2509.13590v3 Announce Type: replace-cross Abstract: The rapid advancement of artificial intelligence (AI) in healthcare imaging has revolutionized diagnostic medicine and clinical decision-makin

Interpretable Traces, Unexpected Outcomes: Investigating the Disconnect in Trace-Based Knowledge Distillation

Model ReleasesDGX agent

arXiv:2505.13792v2 Announce Type: replace-cross Abstract: Recent advances in reasoning-focused Large Language Models (LLMs) have introduced Chain-of-Thought (CoT) traces - intermediate reasoning steps

🚀 Introducing Qwen3.6-Max-Preview, an early preview of our next flagship model Highlights: ⚡️ Improved agentic coding capability over Qwen3…

Model ReleasesDGX agent

🚀 Introducing Qwen3.6-Max-Preview, an early preview of our next flagship model Highlights: ⚡️ Improved agentic coding capability over Qwen3.6-Plus 📖 Stronger world knowledge and instruction following

IPQA: A Benchmark for Core Intent Identification in Personalized Question Answering

Model ReleasesDGX agent

arXiv:2510.23536v2 Announce Type: replace Abstract: Intent identification serves as the foundation for generating appropriate responses in personalized question answering (PQA). However, existing benc

Is this chart lying to me? Automating the detection of misleading visualizations

Model ReleasesDGX agent

arXiv:2508.21675v3 Announce Type: replace Abstract: Misleading visualizations are a potent driver of misinformation on social media and the web. By violating chart design principles, they distort data

I've been a K2.5 superfan since it came out. These new numbers for the next version look incredible. You gotta love competition!

Model ReleasesDGX agent

I've been a K2.5 superfan since it came out. These new numbers for the next version look incredible. You gotta love competition! Meet Kimi K2.6: Advancing Open-Source Coding 🔹Open-source SOTA on HLE w

JFinTEB: Japanese Financial Text Embedding Benchmark

Model ReleasesDGX agent

arXiv:2604.15882v1 Announce Type: cross Abstract: We introduce JFinTEB, the first comprehensive benchmark specifically designed for evaluating Japanese financial text embeddings. Existing embedding be

JumpLoRA: Sparse Adapters for Continual Learning in Large Language Models

Model ReleasesDGX agent

arXiv:2604.16171v1 Announce Type: cross Abstract: Adapter-based methods have become a cost-effective approach to continual learning (CL) for Large Language Models (LLMs), by sequentially learning a lo

Kimi K2.6 just dropped. And it crushed Claude Opus 4.6 on SWE-Bench Pro. Kimi K2.6: 58.6 GPT-5.4 xhigh: 57.7 Gemini 3.1 Pro: 54.2 Claude Opu…

Model ReleasesDGX agent

Kimi K2.6 just dropped. And it crushed Claude Opus 4.6 on SWE-Bench Pro. Kimi K2.6: 58.6 GPT-5.4 xhigh: 57.7 Gemini 3.1 Pro: 54.2 Claude Opus 4.6: 53.4 An open source Chinese model is now #1 on agenti

Kimi K2.6 now in OpenCode — Go included

Model ReleasesDGX agent

Kimi K2.6, an AI model developed by Moonshot, has been released on OpenCode with support for the Go programming language. This release expands the model's capabilities to handle Go code generation, an

Kimi K2.6 raises the bar for open-source models. 🦙 available on Ollama's cloud! Try it with OpenClaw: ollama launch openclaw --model kimi-k…

Model ReleasesDGX agent

Kimi K2.6 raises the bar for open-source models. 🦙 available on Ollama's cloud! Try it with OpenClaw: ollama launch openclaw --model kimi-k2.6:cloud Try it with Hermes Agent: ollama launch hermes --mo

Kimi wires up auth + database + backend in one pass One prompt gets you user registration, login, database, booking systems, admin dashboard…

Model ReleasesDGX agent

Kimi wires up auth + database + backend in one pass One prompt gets you user registration, login, database, booking systems, admin dashboards - wired up and deployed. No separate 'now build the backen

KWBench: Measuring Unprompted Problem Recognition in Knowledge Work

Model ReleasesDGX agent

arXiv:2604.15760v1 Announce Type: new Abstract: We introduce the first version of KWBench (Knowledge Work Bench), a benchmark for unprompted problem recognition in large language models: can an LLM id

LaMSUM: Amplifying Voices Against Harassment through LLM Guided Extractive Summarization of User Incident Reports

Model ReleasesDGX agent

arXiv:2406.15809v5 Announce Type: replace Abstract: Citizen reporting platforms help the public and authorities stay informed about sexual harassment incidents. However, the high volume of data shared

Last week, Anthropic dropped the coolest 'AI isn't just chat' product. Claude Design lets you describe what you want to Claude and it return…

Model ReleasesDGX agent

Last week, Anthropic dropped the coolest 'AI isn't just chat' product. Claude Design lets you describe what you want to Claude and it returns prototypes, slides, and one-pagers by just chatting. You c

life when you discover an open-source model that runs 300 parallel agents, executes for 12+ hours straight, beats GPT-5.4 and opus 4.6 on mu…

Model ReleasesDGX agent

life when you discover an open-source model that runs 300 parallel agents, executes for 12+ hours straight, beats GPT-5.4 and opus 4.6 on multiple benchmarks... and the weights are on huggingface Medi

LinuxArena: A Control Setting for AI Agents in Live Production Software Environments

Model ReleasesDGX agent

arXiv:2604.15384v1 Announce Type: cross Abstract: We introduce LinuxArena, a control setting in which agents operate directly on live, multi-service production environments. LinuxArena contains 20 env

LLM attribution analysis across different fine-tuning strategies and model scales for automated code compliance

Model ReleasesDGX agent

arXiv:2604.15589v1 Announce Type: cross Abstract: Existing research on large language models (LLMs) for automated code compliance has primarily focused on performance, treating the models as black box

LLMs Corrupt Your Documents When You Delegate

Model ReleasesDGX agent

arXiv:2604.15597v1 Announce Type: new Abstract: Large Language Models (LLMs) are poised to disrupt knowledge work, with the emergence of delegated work as a new interaction paradigm (e.g., vibe coding

LLMSniffer: Detecting LLM-Generated Code via GraphCodeBERT and Supervised Contrastive Learning

Model ReleasesDGX agent

arXiv:2604.16058v1 Announce Type: cross Abstract: The rapid proliferation of Large Language Models (LLMs) in software development has made distinguishing AI-generated code from human-written code a cr

Making Image Editing Easier via Adaptive Task Reformulation with Agentic Executions

Model ReleasesDGX agent

arXiv:2604.15917v1 Announce Type: new Abstract: Instruction guided image editing has advanced substantially with recent generative models, yet it still fails to produce reliable results across many se

Mamba-SSM with LLM Reasoning for Feature Selection: Faithfulness-Aware Biomarker Discovery

Model ReleasesDGX agent

arXiv:2604.14334v2 Announce Type: replace-cross Abstract: Gradient saliency from deep sequence models surfaces candidate biomarkers efficiently, but the resulting gene lists can be contaminated by tis

MEDLEY-BENCH: Scale Buys Evaluation but Not Control in AI Metacognition

Model ReleasesDGX agent

arXiv:2604.16009v1 Announce Type: new Abstract: Metacognition, the ability to monitor and regulate one's own reasoning, remains under-evaluated in AI benchmarking. We introduce MEDLEY-BENCH, a benchma

MemEvoBench: Benchmarking Memory MisEvolution in LLM Agents

Model ReleasesDGX agent

arXiv:2604.15774v1 Announce Type: new Abstract: Equipping Large Language Models (LLMs) with persistent memory enhances interaction continuity and personalization but introduces new safety risks. Speci

memory will be the great lock in and everyone knows it, so are rushing to get there first memory should be open!

Model ReleasesDGX agent

memory will be the great lock in and everyone knows it, so are rushing to get there first memory should be open! Last week, we released a preview of memories in Codex. Today, we’re expanding the exper

Mind DeepResearch Technical Report

Model ReleasesDGX agent

arXiv:2604.14518v2 Announce Type: replace Abstract: We present Mind DeepResearch (MindDR), an efficient multi-agent deep research framework that achieves leading performance with only ~30B-parameter m

← Previous
1…333334335336337…372
Next →