AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,563 results
Model Releases

Hypergraph and Latent ODE Learning for Multimodal Root Cause Localization in Microservices

DGX agent

arXiv:2605.00351v1 Announce Type: new Abstract: Root cause localization in cloud native microservice systems requires modeling complex service dependencies, irregular temporal dynamics, and heterogene

model-releasesarxiv-cs-lg
4 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

I had Claude Code for web build me this WebAssembly playground for trying out the new Redis array commands https://tools.simonwillison.net/r…

DGX agent

I had Claude Code for web build me this WebAssembly playground for trying out the new Redis array commands https://tools.simonwillison.net/redis-array More notes here: https://simonwillison.net/2026/M

model-releasessimon-willison--x
4 May 2026
Model Releases

I think the fact that GPT-4o and Llama 3.3-80B did no significant harm is just as important as whether AI helped. If older (less accurate & …

DGX agent

I think the fact that GPT-4o and Llama 3.3-80B did no significant harm is just as important as whether AI helped. If older (less accurate & more sycophantic) chatbots essentially did nothing for peopl

model-releasesethan-mollick--x
4 May 2026
Model Releases

InterChart: Benchmarking Visual Reasoning Across Decomposed and Distributed Chart Information

DGX agent

arXiv:2508.07630v2 Announce Type: replace Abstract: We introduce InterChart, a diagnostic benchmark that evaluates how well vision-language models (VLMs) reason across multiple related charts, a task

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Introducing nanowhale 🐳! A tiny DeepSeek model fully pretrained by an agent. Inspired by @karpathy's nanochat, we gave ml-intern the task o…

DGX agent

Introducing nanowhale 🐳! A tiny DeepSeek model fully pretrained by an agent. Inspired by @karpathy's nanochat, we gave ml-intern the task of training a tiny MoE with all the architectural advancements

model-releasesclem-delangue--x
4 May 2026
Model Releases

Introducing WARM-VR: Benchmark Dataset for Multimodal Wearable Affect Recognition in Virtual Reality

DGX agent

arXiv:2605.00184v1 Announce Type: new Abstract: With the growing integration of human-computer interaction into everyday life, advances in machine learning have enabled systems to better perceive and

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Jailbreaking Vision-Language Models Through the Visual Modality

DGX agent

arXiv:2605.00583v1 Announce Type: new Abstract: The visual modality of vision-language models (VLMs) is an underexplored attack surface for bypassing safety alignment. We introduce four jailbreak atta

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Jailbroken Frontier Models Retain Their Capabilities

DGX agent

arXiv:2605.00267v1 Announce Type: new Abstract: As language model safeguards become more robust, attackers are pushed toward developing increasingly complex jailbreaks. Prior work has found that this

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Knowing When to Defer: Selective Prediction for Responsible Knowledge Tracing

DGX agent

arXiv:2509.21514v3 Announce Type: replace-cross Abstract: Research on Knowledge Tracing (KT) models traditionally focuses on improving predictive accuracy. However, responsible real-world deployment r

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Learning from Compressed CT: Feature Attention Style Transfer and Structured Factorized Projections for Resource-Efficient Medical Image Analysis

DGX agent

arXiv:2605.00448v1 Announce Type: new Abstract: The deployment of artificial intelligence in medical imaging is hindered by high computational complexity and resource-intensive processing of volumetri

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Learning from Supervision with Semantic and Episodic Memory: A Reflective Approach to Agent Adaptation

DGX agent

arXiv:2510.19897v2 Announce Type: replace Abstract: We investigate how agents built on pretrained large language models (LLMs) can learn target classification functions from labeled examples without p

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Learning from the Unseen: Generative Data Augmentation for Geometric-Semantic Accident Anticipation

DGX agent

arXiv:2605.00051v1 Announce Type: new Abstract: Anticipating traffic accidents is a critical yet unresolved problem for autonomous driving, hindered by the inherent complexity of modeling interactions

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Learning Locally, Revising Globally: Global Reviser for Federated Learning with Noisy Labels

DGX agent

arXiv:2412.00452v2 Announce Type: replace-cross Abstract: Conventioanl federated learning (FL) heavily depends on high-quality labels, which are often impractical in the real world, leading to the fed

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Lightweight Domain Adaptation of a Large Language Model for Legal Assistance in the Indian Context

DGX agent

arXiv:2505.22003v2 Announce Type: replace Abstract: In India, access to legal assistance for the general public has been observed to have a critical gap, as many citizens are not able to take full adv

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

LLM-Oriented Information Retrieval: A Denoising-First Perspective

DGX agent

arXiv:2605.00505v1 Announce Type: cross Abstract: Modern information retrieval (IR) is no longer consumed primarily by humans but increasingly by large language models (LLMs) via retrieval-augmented g

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

LWiAI Podcast #243 - GPT 5.5, DeepSeek V4, AI safety sabotage

DGX agent

This podcast episode from Last Week in AI discusses recent developments in large language models, including updates on GPT 5.5 and DeepSeek V4, while also covering concerns about potential sabotage or

model-releaseslast-week-in-ai
4 May 2026
Model Releases

M-CaStLe: Uncovering Local Causal Structures in Multivariate Space-Time Gridded Data

DGX agent

arXiv:2605.00398v1 Announce Type: new Abstract: Causal graph discovery for space-time systems is challenging in high-dimensional gridded data, which often has many more grid cells than temporal observ

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Make Your LVLM KV Cache More Lightweight

DGX agent

arXiv:2605.00789v1 Announce Type: new Abstract: Key-Value (KV) cache has become a de facto component of modern Large Vision-Language Models (LVLMs) for inference. While it enhances decoding efficiency

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

May 5 is the GPT-5.5 launch celebration in San Francisco and the Claude Finance Briefing in New York. Real opposite valence events on opposi…

DGX agent

I cannot provide a summary for this entry as the post content appears incomplete or corrupted in the provided information. The title is cut off mid-sentence and doesn't clearly convey the full topic.

model-releasesethan-mollick--x
4 May 2026
Model Releases

MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

DGX agent

arXiv:2510.17281v5 Announce Type: replace Abstract: Scaling up data, parameters, and test-time computation has been the mainstream methods to improve LLM systems (LLMsys), but their upper bounds are a

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Minimizing Human Intervention in Online Classification

DGX agent

arXiv:2510.23557v2 Announce Type: replace-cross Abstract: Training or fine-tuning large language model (LLM)-based systems often requires costly human feedback, yet there is limited understanding of h

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

ML-Agent: Reinforcing LLM Agents for Autonomous Machine Learning Engineering

DGX agent

arXiv:2505.23723v2 Announce Type: replace Abstract: The emergence of large language model (LLM)-based agents has significantly advanced the development of autonomous machine learning (ML) engineering.

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

ML-Bench&Guard: Policy-Grounded Multilingual Safety Benchmark and Guardrail for Large Language Models

DGX agent

arXiv:2605.00689v1 Announce Type: new Abstract: As Large Language Models (LLMs) are increasingly deployed in cross-linguistic contexts, ensuring safety in diverse regulatory and cultural environments

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

MoDAl: Self-Supervised Neural Modality Discovery via Decorrelation for Speech Neuroprosthesis

DGX agent

arXiv:2605.00025v1 Announce Type: cross Abstract: Speech neuroprosthesis systems decode intended speech from neural activity in the absence of audible output, offering a path to restoring communicatio

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Multi-frame Restoration for High-rate Lissajous Confocal Laser Endomicroscopy

DGX agent

arXiv:2605.00527v1 Announce Type: cross Abstract: Lissajous confocal laser endomicroscopy (CLE) is a promising solution for high speed in vivo optical biopsy for handheld scenarios. However, Lissajous

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

NorBERTo: A ModernBERT Model Trained for Portuguese with 331 Billion Tokens Corpus

DGX agent

arXiv:2605.00086v1 Announce Type: new Abstract: High-quality corpora are essential for advancing Natural Language Processing (NLP) in Portuguese. Building on previous encoder-only models such as BERTi

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

🤯 Ollama now supports Claude Desktop via Claude’s built-in third party inference. ollama launch claude-desktop This allows all models from …

DGX agent

🤯 Ollama now supports Claude Desktop via Claude’s built-in third party inference. ollama launch claude-desktop This allows all models from Ollama's Cloud to be used across Claude Cowork and Claude Cod

model-releasesollama--x
4 May 2026
Model Releases

OTSS: Output-Targeted Soft Segmentation for Contextual Decision-Weight Learning

DGX agent

arXiv:2605.00193v1 Announce Type: new Abstract: Many machine learning systems make constrained decisions by optimizing factorized objectives, but the context-specific objective is often treated as fix

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Paired-CSLiDAR: Height-Stratified Registration for Cross-Source Aerial-Ground LiDAR Pose Refinement

DGX agent

arXiv:2605.00634v1 Announce Type: cross Abstract: We introduce Paired-CSLiDAR (CSLiDAR), a cross-source aerial-ground LiDAR benchmark for single-scan pose refinement: refining a ground-scan pose withi

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

PEACE: Cross-modal Enhanced Pediatric-Adult ECG Alignment for Robust Pediatric Diagnosis

DGX agent

arXiv:2605.00647v1 Announce Type: new Abstract: Automated pediatric electrocardiogram (ECG) diagnosis remains challenging because models trained predominantly on adult data suffer from substantial cro

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Peek2: Regex-free Byte-level Byte-Pair Encoding Pretokenizer for LLM Inference on Edge Devices

DGX agent

arXiv:2601.05833v2 Announce Type: replace Abstract: Pretokenization is a crucial, sequential pass in Byte-level BPE tokenizers, yet little work has been done to optimize it for edge-side inference. Ou

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs

DGX agent

arXiv:2605.00814v1 Announce Type: new Abstract: While autoregressive Large Vision-Language Models (LVLMs) demonstrate remarkable proficiency in multimodal tasks, they face a 'Visual Signal Dilution' p

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

PILIR: Physics-Informed Local Implicit Representation

DGX agent

arXiv:2605.00385v1 Announce Type: new Abstract: Physics-Informed Neural Networks have become a powerful mesh-free method for solving partial differential equations, but their performance is often limi

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Pinecone Launches First Serverless Region in Asia with New Singapore Cloud Region, Bringing the Knowledge Infrastructure for AI to the Asia-Pacific Market

DGX agent

Pinecone announced the launch of its first serverless region in Asia, specifically in Singapore, expanding its vector database infrastructure to the Asia-Pacific market. This new cloud region enables

model-releasespinecone
4 May 2026
Model Releases

Poems that ChatGPT, Claude, and Gemini all seem to 'like' when you ask for poetry related to being/making LLMs: Rilke's 'Archaic Torso of Ap…

DGX agent

Poems that ChatGPT, Claude, and Gemini all seem to 'like' when you ask for poetry related to being/making LLMs: Rilke's 'Archaic Torso of Apollo' Stevens' 'Idea of Order at Key West' Borges's 'The Gol

model-releasesethan-mollick--x
4 May 2026
Model Releases

Preference Goal Tuning: Post-Training as Latent Control for Frozen Policies

DGX agent

arXiv:2412.02125v2 Announce Type: replace-cross Abstract: Goal-conditioned policies enable decision-making models to execute diverse behaviors based on specified goals, yet their downstream performanc

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Probing Multimodal Large Language Models on Cognitive Biases in Chinese Short-Video Misinformation

DGX agent

arXiv:2601.06600v2 Announce Type: replace Abstract: Short-video platforms have become major channels for misinformation, where deceptive claims frequently leverage visual experiments and social cues.

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Putting HUMANS first: Efficient LAM Evaluation with Human Preference Alignment

DGX agent

arXiv:2605.00022v1 Announce Type: new Abstract: The rapid proliferation of large audio models (LAMs) demands efficient approaches for model comparison, yet comprehensive benchmarks are costly. To fill

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Real-Time Frame- and Event-based Object Detection with Spiking Neural Networks on Edge Neuromorphic Hardware: Design, Deployment and Benchmark

DGX agent

arXiv:2605.00146v1 Announce Type: new Abstract: Real-time object detection on energy-constrained platforms is critical for applications such as UAV-based inspection, autonomous navigation, and mobile

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Reasoning-Intensive Regression

DGX agent

arXiv:2508.21762v3 Announce Type: replace Abstract: AI researchers and practitioners increasingly apply large language models (LLMs) to what we call reasoning-intensive regression (RiR), i.e., deducin

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Redis Array Playground

DGX agent

Tool: Redis Array Playground Salvatore Sanfilippo submitted a PR adding a new data type - arrays - to Redis. The new commands are ARCOUNT, ARDEL, ARDELRANGE, ARGET, ARGETRANGE, ARGREP, ARINFO, ARINSER

model-releasessimon-willison
4 May 2026
Model Releases

Remote SAMsing: From Segment Anything to Segment Everything

DGX agent

arXiv:2605.00256v1 Announce Type: new Abstract: SAM2 produces high-quality zero-shot segmentation on natural images, but applying it to large remote sensing scenes exposes two problems: (1) its mask g

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

RETO: A Rotary-Enhanced Transformer Operator for High-Fidelity Prediction of Automotive Aerodynamics

DGX agent

arXiv:2605.00062v1 Announce Type: cross Abstract: Rapid aerodynamic evaluation is crucial for modern vehicle design, yet existing neural operators struggle to capture intricate spatial correlations. W

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Retrieval-Augmented Reasoning for Chartered Accountancy

DGX agent

arXiv:2605.00257v1 Announce Type: new Abstract: The inception of Large Language Models (LLMs) has catalyzed AI adoption in the finance sector, yet their reliability in complex, jurisdiction-specific t

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

RSAT: Structured Attribution Makes Small Language Models Faithful Table Reasoners

DGX agent

arXiv:2605.00199v1 Announce Type: new Abstract: When a language model answers a table question, users have no way to verify which cells informed which reasoning steps. We introduce RSAT, a method that

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

RTPrune: Reading-Twice Inspired Token Pruning for Efficient DeepSeek-OCR Inference

DGX agent

arXiv:2605.00392v1 Announce Type: new Abstract: DeepSeek-OCR leverages visual-text compression to reduce long-text processing costs and accelerate inference, yet visual tokens remain prone to redundan

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

SC-Taxo: Hierarchical Taxonomy Generation under Semantic Consistency Constraints using Large Language Models

DGX agent

arXiv:2605.00620v1 Announce Type: new Abstract: Scientific literature is expanding at an unprecedented pace, making it increasingly challenging to efficiently organize and access domain knowledge. A h

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Scalable Context-Aware Graph Attention for Unsupervised Anomaly Detection in Large-Scale Mobile Networks

DGX agent

arXiv:2605.00482v1 Announce Type: new Abstract: Mobile network operators must monitor thousands of heterogeneous network elements across the radio access network and the packet core, each exposing hig

model-releasesarxiv-cs-lg
4 May 2026
← Previous
1…365366367368369…471
Next →