AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,805 results
Model Releases

Adapting Multilingual Embedding Models to Turkish via Cross-Lingual Tokenizer Surgery and Offline Distillation

DGX agent

arXiv:2605.29992v1 Announce Type: new Abstract: Sentence embeddings are a foundational component for semantic search, clustering, classification, and retrieval-augmented generation. This paper present

model-releasesarxiv-cs-cl
29 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

AfriScience-MT: Towards Decolonizing Science in Africa through Text Translation

DGX agent

arXiv:2605.29741v1 Announce Type: new Abstract: The dominance of colonial languages in African education and scientific communication limits how hundreds of millions of speakers of African languages a

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

AgentCVR: Active Multi-Agent Cross-Video Reasoning via Script-Simulated Reinforcement Learning

DGX agent

arXiv:2605.29643v1 Announce Type: new Abstract: Cross-Video Reasoning (CVR) has emerged as a critical frontier in multimodal intelligence, requiring models to retrieve, align, and aggregate evidence d

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

AgentDoG 1.5: A Lightweight and Scalable Alignment Framework for AI Agent Safety and Security

DGX agent

arXiv:2605.29801v1 Announce Type: new Abstract: Modern open-world agents such as OpenClaw exhibit powerful cross-environment execution capabilities yet introduce broad new safety risk sources. Meanwhi

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

AgentDropoutV2: Optimizing Information Flow in Multi-Agent Systems via Test-Time Rectify-or-Reject Pruning

DGX agent

arXiv:2602.23258v2 Announce Type: replace Abstract: While Multi-Agent Systems (MAS) excel in complex reasoning, they suffer from the cascading impact of erroneous information from individual agents. C

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

AI can give researchers the freedom to pursue “crazier” ideas. For Terence Tao, AI creates more room to experiment, test unexpected paths, a…

DGX agent

AI tools are enabling researchers, including renowned mathematician Terence Tao, to explore unconventional and high-risk ideas by handling routine computational tasks and verification work. This techn

model-releasesopenai--x
29 May 2026
Model Releases

AI startup Shift launches a free home cleaning service in NYC to record first-person video with a camera-equipped cap and use it to train robots (Robert Hart/The Verge)

DGX agent

Robert Hart / The Verge: AI startup Shift launches a free home cleaning service in NYC to record first-person video with a camera-equipped cap and use it to train robots — Shift says a ‘magic hat’ wi

model-releasestechmeme
29 May 2026
Model Releases

Aligned but Fragile: Enhancing LLM Safety Robustness via Zeroth-Order Optimization

DGX agent

arXiv:2605.29396v1 Announce Type: new Abstract: Safety alignment for large language models (LLMs) aims to reduce harmful or unsafe behavior while preserving general utility. However, recent findings r

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Alignment-Guided Score Matching for Text-to-Image Alignment in Diffusion Models

DGX agent

arXiv:2605.30038v1 Announce Type: cross Abstract: Diffusion models generate highly realistic images but often struggle with precise text-image alignment. While recent post-training methods improve ali

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

AlignVid: Training-Free Attention Scaling for Semantic Fidelity in Text-Guided Image-to-Video Generation

DGX agent

arXiv:2512.01334v2 Announce Type: replace Abstract: Text-guided image-to-video generation has made substantial progress, yet it still struggles to execute text-specified edits that require substantial

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

AlloyDB Hot Standby: Faster failovers, consistent performance

DGX agent

AlloyDB for PostgreSQL is a fully managed, PostgreSQL-compatible database service designed for the most demanding enterprise workloads. It combines the best of PostgreSQL with the power of Google, del

model-releasesgoogle-cloud-ai
29 May 2026
Model Releases

AMDP: Asynchronous Multi-Directional Pipeline Parallelism for Large-Scale Models Training

DGX agent

arXiv:2605.29664v1 Announce Type: cross Abstract: Pipeline parallelism is essential for large-scale model training, but existing asynchronous approaches often degrade convergence due to parameter mism

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

An End-to-End PyTorch Interface for Differentiable PDE Solvers: A RANS Model-Correction Study

DGX agent

arXiv:2605.28858v1 Announce Type: cross Abstract: This work presents an end-to-end strategy for solving inverse problems constrained by Partial Differential Equations within a fully differentiable Mac

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

Another proof point for the open-weights thesis. From @RampLabs: 'If we built this again, we'd lean more on open-weight models.' Ramp pointe…

DGX agent

Another proof point for the open-weights thesis. From @RampLabs: 'If we built this again, we'd lean more on open-weight models.' Ramp pointed 10K agents at their own backend. Kimi K2.6 and DeepSeek V4

model-releasesfireworks-ai--x
29 May 2026
Model Releases

Anthropic's run-rate revenue hits $47 billion

DGX agent

The most interesting thing about Anthropic's 65B Series H announcement is this line (emphasis mine): Since our Series G in February, adoption has continued to grow across global enterprise customers,

model-releasessimon-willison
29 May 2026
Model Releases

Apertus LLM Family Expansion via Distillation and Quantization

DGX agent

arXiv:2605.29128v1 Announce Type: new Abstract: The wide adoption of LLMs has led to their use in great variety of applications and scenarios, such as chatbot assistants and data annotation, creating

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

Architecture-Sensitive Supervised Fine-Tuning for Screen-Conditioned Action Prediction: A PiSAR Benchmark

DGX agent

arXiv:2605.29400v1 Announce Type: new Abstract: We benchmark three supervised fine-tuned models against frontier zero-shot baselines on a 661-row held-out slice of PiSAR (Persona, intent, Screen, Acti

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Are LLMs Socially Adaptive? Contrasting Belief Evolution in Large Language Models and Humans

DGX agent

arXiv:2410.10398v3 Announce Type: replace-cross Abstract: As large language models (LLMs) increasingly engage in complex social interactions, ensuring that their behaviors align with human ethical pri

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

AtomWorld: A Benchmark for Evaluating Spatial Reasoning in Large Language Models on Crystalline Materials

DGX agent

arXiv:2510.04704v4 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown promising potential in scientific research, enabling tasks ranging from knowledge retrieval to propert

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

AttuneBench: A Conversation-Based Benchmark for LLM Emotional Intelligence

DGX agent

arXiv:2605.21739v2 Announce Type: replace Abstract: Emotional intelligence (EI), the ability to perceive, understand, and respond appropriately to others' emotional states, is central to human communi

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Audio Deepfake Detection with Half-Truth Localisation Using Cross-Attentive Feature Fusion

DGX agent

arXiv:2605.29531v1 Announce Type: cross Abstract: Audio deepfake detection is well-studied as a binary problem, but partially manipulated speech, where a short synthesised segment is spliced into an o

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

Auditing Training-Free 3D Shape Retrieval with Diffused Geodesic Moments

DGX agent

arXiv:2605.29004v1 Announce Type: new Abstract: Reported retrieval scores for training-free shape descriptors conflate local signal design, normalization, aggregation, codebook fitting, and metric cho

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

AutoSizer: Automatic Sizing of Analog and Mixed-Signal Circuits via Large Language Model (LLM) Agents

DGX agent

arXiv:2602.02849v2 Announce Type: replace Abstract: The design of Analog and Mixed-Signal (AMS) integrated circuits remains heavily reliant on expert knowledge, with transistor sizing a major bottlene

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Balancing Multimodal Learning through Label Space Reshaping

DGX agent

arXiv:2605.28869v1 Announce Type: cross Abstract: Multimodal learning often suffers from modality imbalance, where modalities that converge faster dominate optimization while others remain undertraine

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Bandit Algorithms for Deep Brain Stimulation

DGX agent

arXiv:2601.12699v2 Announce Type: replace Abstract: Deep Brain Stimulation (DBS) is an effective treatment for Parkinson's disease, but conventional fixed-parameter stimulation can reduce battery life

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

Battery-Sim-Agent: Leveraging LLM-Agent for Inverse Battery Parameter Estimation

DGX agent

arXiv:2605.29560v1 Announce Type: new Abstract: Parameterizing high-fidelity 'digital twins' of batteries is a critical yet challenging inverse problem that hinders the pace of battery innovation. Pre

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

'Be My Cheese?': Cultural Nuance Benchmarking for Machine Translation in Multilingual LLMs

DGX agent

arXiv:2602.04729v2 Announce Type: replace Abstract: We present a large-scale human evaluation benchmark for assessing cultural localisation in machine translation produced by state-of-the-art multilin

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

Benchmarking Large Vision-Language Models on CFMME: A Comprehensive Chinese Financial Multimodal Evaluation Dataset

DGX agent

arXiv:2605.29462v1 Announce Type: cross Abstract: The emergence of Large Vision-Language Models (LVLMs) has substantially expanded model capabilities beyond text-only understanding, enabling unified i

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting

DGX agent

arXiv:2509.23571v3 Announce Type: replace-cross Abstract: As cyber threats continue to grow in scale and sophistication, blue team defenders increasingly require advanced tools to proactively detect a

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Benchmarking Open-Source Safety Guard Models: A Comprehensive Evaluation

DGX agent

arXiv:2605.28830v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are increasingly deployed in safety-critical applications, robust content moderation becomes essential. We present a c

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Benchmarking Positional Encoding Strategies for Transformer-Based EEG Foundation Models

DGX agent

arXiv:2605.29754v1 Announce Type: new Abstract: Electroencephalography (EEG) is a widely used non-invasive technique for measuring brain activity in brain-computer interface (BCI) applications. Superv

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Benchmarking Single-Factor Physical Video-to-Audio Generation

DGX agent

arXiv:2605.30339v1 Announce Type: new Abstract: Generative video-to-audio (V2A) models produce highly plausible soundtracks, but it remains unclear whether they capture the underlying physical process

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

BenchTrace: A Benchmark for Testing Reflection Ability and Controlled Evolution in LLM Agents

DGX agent

arXiv:2605.29225v1 Announce Type: new Abstract: Self-evolving agents improve over time by reflecting on past failures, but existing evaluation is limited in two ways: it measures only task scores, lea

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Beyond English and Evasion: A Human-Annotated Multi-Domain Benchmark for High-Stakes LLM Safety Evaluation in Chinese

DGX agent

arXiv:2605.29667v1 Announce Type: new Abstract: When Large Language Models (LLMs) are deployed in Chinese-language settings, a troubling pattern emerges: safety systems that work well in English break

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

Beyond Recall: Behavioral Specification as an Interpretive Layer for AI Personalization

DGX agent

arXiv:2605.28969v1 Announce Type: cross Abstract: If an AI agent makes decisions on a person's behalf, those decisions must align with its user. We introduce representational accuracy to measure how f

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

BioRefusalAudit: Auditing Biosecurity Refusal Depth Using General and Domain-Fine-Tuned Sparse Autoencoders

DGX agent

arXiv:2605.30162v1 Announce Type: new Abstract: Biosecurity evaluations of language models typically ask whether models produce hazardous output. This paper asks a complementary question: when a model

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Boston Children’s uses AI to unlock new diagnoses

DGX agent

Boston Children's Hospital has implemented AI technology to improve diagnostic accuracy and identify rare or complex medical conditions in pediatric patients that might otherwise go undiagnosed. The a

model-releasesopenai
29 May 2026
Model Releases

BrahmicTokenizer-131K: An Indic-Capable Drop-In Replacement for o200k_base

DGX agent

arXiv:2605.29379v1 Announce Type: new Abstract: We present BrahmicTokenizer-131K, a 131,072-vocabulary byte-level BPE tokenizer that closes the Brahmic compression gap at the 131K-vocabulary class whi

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

Brain-IT-VQA: From Brain Signals to Answers

DGX agent

arXiv:2605.29588v1 Announce Type: cross Abstract: Decoding visual content from fMRI signals recorded while a person views images, and specifically answering questions about the seen images, is a long-

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Bridging the Semantic Gap for Categorical Data Clustering via Large Language Models

DGX agent

arXiv:2601.01162v3 Announce Type: replace-cross Abstract: Qualitative data are widespread in domains such as healthcare, marketing, and bioinformatics, where clustering offers a fundamental tool for p

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Building and Road Recognition in Dense Urban Informal Settlements: A Dataset and Benchmark

DGX agent

arXiv:2605.29856v1 Announce Type: new Abstract: As a widespread form of informal settlements, urban villages present significant challenges for sustainable urban development and governance. Precise ma

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

BullingerDB: A Dataset for Handwritten Text Recognition and Writer Retrieval

DGX agent

arXiv:2605.30235v1 Announce Type: new Abstract: We present BullingerDB, a large-scale benchmark dataset for historical document analysis based on the correspondence of Heinrich Bullinger (1504-1575).

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

CalArena: A Large-Scale Post-Hoc Calibration Benchmark

DGX agent

arXiv:2605.30188v1 Announce Type: cross Abstract: Reliable probability estimates are critical in many machine learning applications, yet modern classifiers are often poorly calibrated. Post-hoc calibr

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Can AI Weather Models Predict Beyond Two Weeks? A Quantitative Benchmark and Analysis of Long Rollouts

DGX agent

arXiv:2605.30184v1 Announce Type: new Abstract: While AI weather models excel at short-to-medium range forecasts (up to 15 days), they frequently suffer from ill-defined 'instabilities' when rolled ou

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

Casual as an Anchor: Resolving Supervision Misalignment in Formality Transfer Dataset

DGX agent

arXiv:2605.29365v1 Announce Type: new Abstract: Formality transfer is commonly framed as a symmetric bidirectional task between informal and formal registers. We argue that this framing conceals a sup

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

Certified Causal Defense with Generalizable Robustness

DGX agent

arXiv:2408.15451v3 Announce Type: replace Abstract: While machine learning models have proven effective across various scenarios, it is widely acknowledged that many models are vulnerable to adversari

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

ChatGPT diagnosed 40 million people with a disease that was invented as a joke. Not a real disease. Not a misunderstood disease. A completel…

DGX agent

ChatGPT diagnosed 40 million people with a disease that was invented as a joke. Not a real disease. Not a misunderstood disease. A completely fictional condition with a fake name, fake papers, and fak

model-releasesgary-marcus--x
29 May 2026
Model Releases

Chess-World-Model: A 10M-Game Benchmark for Exact State Tracking from Chess Move Sequences

DGX agent

arXiv:2605.30100v1 Announce Type: new Abstract: World models require state tracking, which is the ability to maintain a correct latent state across action sequences. Existing benchmarks are often synt

model-releasesarxiv-cs-lg
29 May 2026
← Previous
1…244245246247248…476
Next →