AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,280 results
15 Apr 2026

Dual-Modality Anchor-Guided Filtering for Test-time Prompt Tuning

Model ReleasesDGX agent

arXiv:2604.12403v1 Announce Type: new Abstract: Test-Time Prompt Tuning (TPT) adapts vision-language models using augmented views, but its effectiveness is hindered by the challenge of determining whi

EDGE-Shield: Efficient Denoising-staGE Shield for Violative Content Filtering via Scalable Reference-Based Matching

Model ReleasesDGX agent

arXiv:2604.06063v2 Announce Type: replace Abstract: The advent of Text-to-Image generative models poses significant risks of copyright violation and deepfake generation. Since the rapid proliferation

EEG-Based Multimodal Learning via Hyperbolic Mixture-of-Curvature Experts

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.12579v1 Announce Type: new Abstract: Electroencephalography (EEG)-based multimodal learning integrates brain signals with complementary modalities to improve mental state assessment, provid

Efficient Adversarial Training via Criticality-Aware Fine-Tuning

Model ReleasesDGX agent

arXiv:2604.12780v1 Announce Type: cross Abstract: Vision Transformer (ViT) models have achieved remarkable performance across various vision tasks, with scalability being a key advantage when applied

EgoEsportsQA: An Egocentric Video Benchmark for Perception and Reasoning in Esports

Model ReleasesDGX agent

arXiv:2604.12320v1 Announce Type: cross Abstract: While video large language models (Video-LLMs) excel in understanding slow-paced, real-world egocentric videos, their capabilities in high-velocity, i

ELoG-GS: Dual-Branch Gaussian Splatting with Luminance-Guided Enhancement for Extreme Low-light 3D Reconstruction

Model ReleasesDGX agent

arXiv:2604.12592v1 Announce Type: new Abstract: This paper presents our approach to the NTIRE 2026 3D Restoration and Reconstruction Challenge (Track 1), which focuses on reconstructing high-quality 3

Emergent launches Wingman: a personal AI agent for everyone

Model ReleasesDGX agent

Emergent Labs Inc., a vibe coding platform for building production-ready software, today announced the launch of Wingman, a personal, autonomous artificial intelligence agent that helps people manage

Empirical Evaluation of PDF Parsing and Chunking for Financial Question Answering with RAG

Model ReleasesDGX agent

arXiv:2604.12047v1 Announce Type: new Abstract: PDF files are primarily intended for human reading rather than automated processing. In addition, the heterogeneous content of PDFs, such as text, table

Evaluating Differential Privacy Against Membership Inference in Federated Learning: Insights from the NIST Genomics Red Team Challenge

Model ReleasesDGX agent

arXiv:2604.12737v1 Announce Type: cross Abstract: While Federated Learning (FL) mitigates direct data exposure, the resulting trained models remain susceptible to membership inference attacks (MIAs).

Evaluating LLM-Generated ACSL Annotations for Formal Verification

Model ReleasesDGX agent

arXiv:2602.13851v3 Announce Type: replace-cross Abstract: Formal specifications are crucial for building verifiable and dependable software systems, yet generating accurate and verifiable specificatio

Evaluating Relational Reasoning in LLMs with REL

Model ReleasesDGX agent

arXiv:2604.12176v1 Announce Type: new Abstract: Relational reasoning is the ability to infer relations that jointly bind multiple entities, attributes, or variables. This ability is central to scienti

FABLE: Fine-grained Fact Anchoring for Unstructured Model Editing

Model ReleasesDGX agent

arXiv:2604.12559v1 Announce Type: new Abstract: Unstructured model editing aims to update models with real-world text, yet existing methods often memorize text holistically without reliable fine-grain

FaCT: Faithful Concept Traces for Explaining Neural Network Decisions

Model ReleasesDGX agent

arXiv:2510.25512v2 Announce Type: replace-cross Abstract: Deep networks have shown remarkable performance across a wide range of tasks, yet getting a global concept-level understanding of how they fun

Fast and principled equation discovery from chaos to climate

Model ReleasesDGX agent

arXiv:2604.11929v1 Announce Type: new Abstract: Our ability to predict, control, and ultimately understand complex systems rests on discovering the equations that govern their dynamics. Identifying th

FAST-DIPS: Adjoint-Free Analytic Steps and Hard-Constrained Likelihood Correction for Diffusion-Prior Inverse Problems

Model ReleasesDGX agent

arXiv:2603.01591v2 Announce Type: replace-cross Abstract: Training-free diffusion priors enable inverse-problem solvers without retraining, but for nonlinear forward operators data consistency often r

FeaXDrive: Feasibility-aware Trajectory-Centric Diffusion Planning for End-to-End Autonomous Driving

Model ReleasesDGX agent

arXiv:2604.12656v1 Announce Type: cross Abstract: End-to-end diffusion planning has shown strong potential for autonomous driving, but the physical feasibility of generated trajectories remains insuff

Filtered Reasoning Score: Evaluating Reasoning Quality on a Model's Most-Confident Traces

Model ReleasesDGX agent

arXiv:2604.11996v1 Announce Type: cross Abstract: Should we trust Large Language Models (LLMs) with high accuracy? LLMs achieve high accuracy on reasoning benchmarks, but correctness alone does not re

Fine-tuning Factor Augmented Neural Lasso for Heterogeneous Environments

Model ReleasesDGX agent

arXiv:2604.12288v1 Announce Type: cross Abstract: Fine-tuning is a widely used strategy for adapting pre-trained models to new tasks, yet its methodology and theoretical properties in high-dimensional

FlowBoost Reveals Phase Transitions and Spectral Structure in Finite Free Information Inequalities

Model ReleasesDGX agent

arXiv:2604.11922v1 Announce Type: cross Abstract: Using FlowBoost, a closed-loop deep generative optimization framework for extremal structure discovery, we investigate ell^p-generalizations of the fi

Fragile Preferences: A Deep Dive Into Order Effects in Large Language Models

Model ReleasesDGX agent

arXiv:2506.14092v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed in decision-support systems for high-stakes domains such as hiring and university admissions,

From Imitation to Discrimination: Progressive Curriculum Learning for Robust Web Navigation

Model ReleasesDGX agent

arXiv:2604.12666v1 Announce Type: cross Abstract: Text-based web agents offer computational efficiency for autonomous web navigation, yet developing robust agents remains challenging due to the noisy

From Plan to Action: How Well Do Agents Follow the Plan?

Model ReleasesDGX agent

arXiv:2604.12147v1 Announce Type: cross Abstract: Agents aspire to eliminate the need for task-specific prompt crafting through autonomous reason-act-observe loops. Still, they are commonly instructed

Frontier-Eng: Benchmarking Self-Evolving Agents on Real-World Engineering Tasks with Generative Optimization

Model ReleasesDGX agent

arXiv:2604.12290v1 Announce Type: new Abstract: Current LLM agent benchmarks, which predominantly focus on binary pass/fail tasks such as code generation or search-based question answering, often negl

FRTSearch: Unified Detection and Parameter Inference of Fast Radio Transients using Instance Segmentation

Model ReleasesDGX agent

arXiv:2604.12344v1 Announce Type: cross Abstract: The exponential growth of data from modern radio telescopes presents a significant challenge to traditional single-pulse search algorithms, which are

Fully Homomorphic Encryption on Llama 3 model for privacy preserving LLM inference

Model ReleasesDGX agent

arXiv:2604.12168v1 Announce Type: cross Abstract: The applications of Generative Artificial Intelligence (GenAI) and their intersections with data-driven fields, such as healthcare, finance, transport

Fundus Image-based Glaucoma Screening via Retinal Knowledge-Oriented Dynamic Multi-Level Feature Integration

Model ReleasesDGX agent

arXiv:2604.12351v1 Announce Type: new Abstract: Automated diagnosis based on color fundus photography is essential for large-scale glaucoma screening. However, existing deep learning models are typica

GCA Framework: A Gulf-Grounded Dataset and Agentic Pipeline for Climate Decision Support

Model ReleasesDGX agent

arXiv:2604.12306v1 Announce Type: cross Abstract: Climate decision-making in the Gulf increasingly demands systems that can translate heterogeneous scientific and policy evidence into actionable guida

GeM-EA: A Generative and Meta-learning Enhanced Evolutionary Algorithm for Streaming Data-Driven Optimization

Model ReleasesDGX agent

arXiv:2604.12336v1 Announce Type: cross Abstract: Streaming Data-Driven Optimization (SDDO) problems arise in many applications where data arrive continuously and the optimization environment evolves

Gemini 3.1 Flash TTS

Model ReleasesDGX agent

Gemini 3.1 Flash TTS Google released Gemini 3.1 Flash TTS today, a new text-to-speech model that can be directed using prompts. It's presented via the standard Gemini API using gemini-3.1-flash-tts-pr

Gemini 3.1 Flash TTS is rolling out in Google Vids and is available today in preview via the Gemini API and in @GoogleAIStudio. Whether you’…

Model ReleasesDGX agent

Gemini 3.1 Flash TTS is rolling out in Google Vids and is available today in preview via the Gemini API and in @GoogleAIStudio. Whether you’re creating a pitch deck or recording a passion project, tra

Gemini 3.1 Flash TTS: the next generation of expressive AI speech

Model ReleasesDGX agent

Google DeepMind's Gemini 3.1 Flash TTS is a text-to-speech model delivering improved controllability, expressivity, and quality for developers, enterprises, and everyday users building AI-speech appli

Generative Anonymization in Event Streams

Model ReleasesDGX agent

arXiv:2604.12803v1 Announce Type: new Abstract: Neuromorphic vision sensors offer low latency and high dynamic range, but their deployment in public spaces raises severe data protection concerns. Rece

Generative Refinement Networks for Visual Synthesis

Model ReleasesDGX agent

arXiv:2604.13030v1 Announce Type: new Abstract: While diffusion models dominate the field of visual generation, they are computationally inefficient, applying a uniform computational effort regardless

GF-Score: Certified Class-Conditional Robustness Evaluation with Fairness Guarantees

Model ReleasesDGX agent

arXiv:2604.12757v1 Announce Type: cross Abstract: Adversarial robustness is essential for deploying neural networks in safety-critical applications, yet standard evaluation methods either require expe

Global optimization tailored for graphics processing units: Complete and rigorous search for large-scale nonlinear minimization

Model ReleasesDGX agent

arXiv:2507.01770v4 Announce Type: replace-cross Abstract: This paper introduces a numerical method to enclose the global minimum of a nonlinear function subject to simple bounds on the variables. Usin

GlotOCR Bench: OCR Models Still Struggle Beyond a Handful of Unicode Scripts

Model ReleasesDGX agent

arXiv:2604.12978v1 Announce Type: new Abstract: Optical character recognition (OCR) has advanced rapidly with the rise of vision-language models, yet evaluation has remained concentrated on a small cl

GoodPoint: Learning Constructive Scientific Paper Feedback from Author Responses

Model ReleasesDGX agent

arXiv:2604.11924v1 Announce Type: new Abstract: While LLMs hold significant potential to transform scientific research, we advocate for their use to augment and empower researchers rather than to auto

Google adds reusable prompts to Gemini in Chrome

Model ReleasesDGX agent

Google LLC is enhancing the version of its Gemini assistant that is embedded in Chrome with a new time-saving tool called Skills. The capability started rolling out today. It’s accessible on Mac, Wind

Google DeepMind introduces Gemini Robotics-ER 1.6 robotic reasoning model, says it shows significant spatial and physical reasoning improvements over ER 1.5 (Google DeepMind)

Model ReleasesDGX agent

Google DeepMind: Google DeepMind introduces Gemini Robotics-ER 1.6 robotic reasoning model, says it shows significant spatial and physical reasoning improvements over ER 1.5 — For robots to be truly h

Google launches a Gemini AI app on Mac

Model ReleasesDGX agent

Google is launching a new Gemini app on Mac that allows you to interact with the AI assistant without switching windows on your desktop. With the app, you can use the Option + Space shortcut to pull u

Google launches a Gemini Mac app, featuring a keyboard shortcut, screen sharing for better context, image generation with Nano Banana, and more (Abner Li/9to5Google)

Model ReleasesDGX agent

Abner Li / 9to5Google: Google launches a Gemini Mac app, featuring a keyboard shortcut, screen sharing for better context, image generation with Nano Banana, and more — Gemini now has a native Mac app

Google rolls out Gemini 3.1 Flash TTS, a text-to-speech model with support for over 70 languages and audio tags that give developers granular speech control (Matthias Bastian/The Decoder)

Model ReleasesDGX agent

Matthias Bastian / The Decoder: Google rolls out Gemini 3.1 Flash TTS, a text-to-speech model with support for over 70 languages and audio tags that give developers granular speech control — The compa

Great news: the ERNIE editing model is expected to be released by the end of this month

Model ReleasesDGX agent

A Reddit post on r/StableDiffusion announces the anticipated release of an ERNIE image **editing** model from Baidu, complementing the already-available ERNIE-Image-8b generation model from Baidu, whi

GroupKAN: Efficient Kolmogorov-Arnold Networks via Grouped Spline Modeling

Model ReleasesDGX agent

arXiv:2511.05477v2 Announce Type: replace Abstract: Medical image segmentation demands models that achieve high accuracy while maintaining computational efficiency and clinical interpretability. While

Growing Pains: Extensible and Efficient LLM Benchmarking Via Fixed Parameter Calibration

Model ReleasesDGX agent

arXiv:2604.12843v1 Announce Type: new Abstract: The rapid release of both language models and benchmarks makes it increasingly costly to evaluate every model on every dataset. In practice, models are

GTCN-G: A Residual Graph-Temporal Fusion Network for Imbalanced Intrusion Detection

Model ReleasesDGX agent

arXiv:2510.07285v3 Announce Type: replace-cross Abstract: The escalating complexity of network threats and the inherent class imbalance in traffic data present formidable challenges for modern Intrusi

GTPBD-MM: A Global Terraced Parcel and Boundary Dataset with Multi-Modality

Model ReleasesDGX agent

arXiv:2604.12315v1 Announce Type: new Abstract: Agricultural parcel extraction plays an important role in remote sensing-based agricultural monitoring, supporting parcel surveying, precision managemen

Guide to prompting Gemini 3.1 Flash TTS (text-to-speech)

Model ReleasesDGX agent

Today, Gemini 3.1 Flash TTS, our latest text-to-speech model, is available on Google AI Studio and Vertex AI. It delivers precise controllability and expressivity, empowering developers and enterprise

Hard Negative Sample-Augmented DPO Post-Training for Small Language Models

Model ReleasesDGX agent

arXiv:2512.19728v2 Announce Type: replace Abstract: Large language models (LLMs) continue to struggle with mathematical reasoning, and common post-training pipelines often reduce each generated soluti

Have found the same things! Using glm-5 as a daily driver for a lot of things

Model ReleasesDGX agent

Have found the same things! Using glm-5 as a daily driver for a lot of things We've tested new OSS models the moment they're released for a while at Lindy. Inference is our #1 cost by a lot (more than

HazardArena: Evaluating Semantic Safety in Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2604.12447v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models inherit rich world knowledge from vision-language backbones and acquire executable skills via action demonstrations.

How memory can affect collective and cooperative behaviors in an LLM-Based Social Particle Swarm

Model ReleasesDGX agent

arXiv:2604.12250v1 Announce Type: new Abstract: This study examines how model-specific characteristics of Large Language Model (LLM) agents, including internal alignment, shape the effect of memory on

How to make ChatGPT give responses similar to Claude, and not agreeing with everything you say?

Model ReleasesDGX agent

This r/ChatGPT thread addresses the well-documented tendency of ChatGPT to be sycophantic — by default, ChatGPT is trained to be polite, non-confrontational, and agreeable — and contrasts it with Clau

HSG-12M: A Large-Scale Benchmark of Spatial Multigraphs from the Energy Spectra of Non-Hermitian Crystals

Model ReleasesDGX agent

arXiv:2506.08618v4 Announce Type: replace-cross Abstract: AI is transforming scientific research by revealing new ways to understand complex physical systems, but its impact remains constrained by the

I ran the same prompt for a London Estuary accent, a Newcastle accent and an Exeter, Devon accent - all three audio files are now embedded i…

Model ReleasesDGX agent

I ran the same prompt for a London Estuary accent, a Newcastle accent and an Exeter, Devon accent - all three audio files are now embedded in my blog post https://simonwillison.net/2026/Apr/15/gemini-

IAD-Unify: A Region-Grounded Unified Model for Industrial Anomaly Segmentation, Understanding, and Generation

Model ReleasesDGX agent

arXiv:2604.12440v1 Announce Type: cross Abstract: Real-world industrial inspection requires not only localizing defects, but also explaining them in natural language and generating controlled defect e

IDEA: An Interpretable and Editable Decision-Making Framework for LLMs via Verbal-to-Numeric Calibration

Model ReleasesDGX agent

arXiv:2604.12573v1 Announce Type: new Abstract: Large Language Models are increasingly deployed for decision-making, yet their adoption in high-stakes domains remains limited by miscalibrated probabil

Identifying and Mitigating Gender Cues in Academic Recommendation Letters: An Interpretability Case Study

Model ReleasesDGX agent

arXiv:2604.12337v1 Announce Type: new Abstract: Letters of recommendation (LoRs) can carry patterns of implicitly gendered language that can inadvertently influence downstream decisions, e.g. in hirin

Identity as Attractor: Geometric Evidence for Persistent Agent Architecture in LLM Activation Space

Model ReleasesDGX agent

arXiv:2604.12016v1 Announce Type: new Abstract: Large language models map semantically related prompts to similar internal representations -- a phenomenon interpretable as attractor-like dynamics. We

If it can’t parse, it can’t perform. ❌ Reliable document understanding is fundamental to enterprise-grade #AgenticAutomation. We're excited …

Model ReleasesDGX agent

If it can’t parse, it can’t perform. ❌ Reliable document understanding is fundamental to enterprise-grade #AgenticAutomation. We're excited to see the release of ParseBench, a new open-source benchmar

← Previous
1…345346347348349…372
Next →