AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,340 results
30 Jul 2026

Meta-Learned Reward Shaping for Reinforcement Learning from Human Feedback

Model ReleasesDGX agent

arXiv:2607.26094v1 Announce Type: cross Abstract: Reinforcement Learning from Human Feedback (RLHF) is the standard approach for aligning large language models with human preferences, but its quality

Migrate your prompts to new models and optimize them on Amazon Bedrock

Model ReleasesDGX agent

Amazon Bedrock Advanced Prompt Optimization optimizes your prompts for up to 5 models at once and compares original versus optimized performance across quality, latency, and cost. Migrate to a new mod

MiniIO debuts AIStor Memory, the long-term memory AI agents need to scale safely

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Object storage software company MiniIO Inc. says it has cracked the persistent memory problem for artificial intelligence agents with the launch of a new offering called AIStor Memory. Whereas convent

Mixture-of-experts for handwriting trajectory reconstruction from IMU sensors

Model ReleasesDGX agent

arXiv:2607.26708v1 Announce Type: new Abstract: The use of digital pens for online handwriting trajectory reconstruction is a prevalent method for human-computer interaction. In this study, we focus o

ML2B: Benchmarking LLMs on Cross-Lingual ML Pipeline Generation

Model ReleasesDGX agent

arXiv:2509.22768v3 Announce Type: replace Abstract: We introduce ML2B, the first benchmark for evaluating cross-lingual task comprehension in end-to-end ML pipeline generation by large language models

Nanbeige4.2-3B: I'm not impressed

Model ReleasesDGX agent

I've tested Nanbeige-4.2-3B. On paper, the benchmarks promise it blows away Qwen3.5-9B and Gemma4-12B. My goal was to have something very light and fast to replace Qwen3.6-35B (or finetunes thereof) f

Not every page in your PDF needs the same treatment. A scanned cover, a dense table, a clean text page, a figure-heavy diagram.... most pars…

Model ReleasesDGX agent

Not every page in your PDF needs the same treatment. A scanned cover, a dense table, a clean text page, a figure-heavy diagram.... most parsing pipelines throw all of them at the same parser, forcing

OmegaUse-OfficeVal: Benchmarking LLM Agents on Long-Horizon Office-Suite Tasks with Economic Grounding

Model ReleasesDGX agent

arXiv:2607.27155v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly expected to assist users in completing tasks. However, existing benchmarks provide limited support

OmniAD: Detect and Understand Industrial Anomaly via Multimodal Reasoning

Model ReleasesDGX agent

arXiv:2505.22039v2 Announce Type: replace Abstract: While anomaly detection has made significant progress, generating detailed analyses that incorporate industrial knowledge remains a challenge. To ad

Online Handwriting Trajectory Reconstruction from Kinematic Sensors using Temporal Convolutional Network

Model ReleasesDGX agent

arXiv:2607.26733v1 Announce Type: new Abstract: Handwriting with digital pens is a common way to facilitate human-computer interaction through the use of Online Handwriting (OH) trajectory reconstruct

OpenAI says it is cutting the price of GPT-5.6 Luna by ~80% and the price of GPT-5.6 Terra by 20% after improving the efficiency of the systems that serve them (Ina Fried/Axios)

Model ReleasesDGX agent

Ina Fried / Axios: OpenAI says it is cutting the price of GPT-5.6 Luna by ~80% and the price of GPT-5.6 Terra by 20% after improving the efficiency of the systems that serve them — OpenAI said Thursda

Our partners at @depthfirstlabs just released dfs-large1, a specialized model built for finding and validating real vulnerabilities in large…

Model ReleasesDGX agent

Our partners at @depthfirstlabs just released dfs-large1, a specialized model built for finding and validating real vulnerabilities in large enterprise codebases. We helped them to scale the training

P.A.I. — Sleek Native Desktop AI Overlayer for Local Ollama Models 🤖⚡

Model ReleasesDGX agent

Greetings Community! 👋 I hope everyone is doing well! I'm Tauhid — Senior EEE student from a Bangladeshi University Today I'd like to share an open-source project I’ve been developing called P.A.I. (P

Parameter-Free Dynamic Regret for Online Convex Optimization under Heavy-Tailed Noise

Model ReleasesDGX agent

arXiv:2607.27073v1 Announce Type: new Abstract: We study online convex optimization (OCO) in non-stationary environments under heavy-tailed noise, where the stochastic gradient oracle admits only a fi

PatchDenoiser: Parameter-efficient multi-scale patch learning and fusion denoiser for Low-dose CT imaging

Model ReleasesDGX agent

arXiv:2602.21987v3 Announce Type: replace Abstract: Low-dose CT images are essential for reducing radiation exposure in cancer screening, pediatric imaging, and longitudinal monitoring protocols, but

Persistence Spheres: a Bi-continuous Linear Representation of Measures for Partial Optimal Transport

Model ReleasesDGX agent

arXiv:2603.15384v2 Announce Type: replace-cross Abstract: We improve and extend persistence spheres, introduced in~ite{pegoraro2025persistence}. Persistence spheres map an integrable measure mu on the

Phoneme- vs. Character-Level Targets and Selective State-Space Models for Intracortical Brain-to-Text

Model ReleasesDGX agent

arXiv:2607.26751v1 Announce Type: new Abstract: State-of-the-art intracortical brain-to-text systems pair a neural-sequence phone decoder with an external language model. Two design axes remain undere

PolyAI launches new real-time voice conversation model to make AI-driven calls more human

Model ReleasesDGX agent

Voice assistant and conversational AI agent developer PolyAI Ltd. today announced the release of Dialog-RSN-1, a voice dialog artificial intelligence model capable of directly perceiving and respondin

Position: Evaluation Scores Are Perishable Knowledge Claims

Model ReleasesDGX agent

arXiv:2607.26191v1 Announce Type: cross Abstract: Evaluation methodologies for language models increasingly combine multiple signals, from automated metrics and LLM-as-judge ratings to human assessmen

Post-Training at the Edge of Detectability: A Game-Theoretic Approach to Fine-Tuning

Model ReleasesDGX agent

arXiv:2607.26358v1 Announce Type: new Abstract: Reinforcement learning (RL) fine-tuning is widely used in language model training to improve model performance on a target task while limiting drift fro

PowerAtlas: Towards Electricity-Computing Co-Scheduling for Power Systems

Model ReleasesDGX agent

arXiv:2607.26710v1 Announce Type: new Abstract: The rapid growth of AI workloads is turning data centers into large-scale, volatile, yet spatiotemporally flexible grid loads, creating an urgent need f

Progressive Multimodal Alignment for Continual Instruction Tuning

Model ReleasesDGX agent

arXiv:2607.26947v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) rely on a projector to align visual representations with the language embedding space, making it central to cro

Projective Graph Residualization: Variation-Allocation Frontiers for Control-Function IV

Model ReleasesDGX agent

arXiv:2606.14636v2 Announce Type: replace Abstract: Control-function instrumental-variable estimators pass an estimated first-stage residual to an outcome model. The residual must retain the latent co

Prosody-driven Jailbreaks in Audio LLMs: A Controlled Study and Mechanistic Analysis

Model ReleasesDGX agent

arXiv:2607.26541v1 Announce Type: cross Abstract: Audio-capable foundation models enable end-to-end spoken interaction, but they also introduce safety risks beyond transcript content. It remains uncle

Quick reminder of what's ok vs not ok with harnesses used for playing ARC-AGI-3: 1. Not okay: harnesses that were custom-made to solve the b…

Model ReleasesDGX agent

Quick reminder of what's ok vs not ok with harnesses used for playing ARC-AGI-3: 1. Not okay: harnesses that were custom-made to solve the benchmark or that contain knowledge about the benchmark forma

Rad-JEPA 3D: Radiology Joint-Embedding Predictive Model for 3D Computed Tomography

Model ReleasesDGX agent

arXiv:2607.26196v1 Announce Type: new Abstract: Self-supervised pretraining is central to 3D medical image analysis, where unlabeled CT volumes are abundant but expert annotations are scarce. Yet exis

RAGuard: A Layered Defense Framework for Retrieval-Augmented Generation Systems Against Data Poisoning

Model ReleasesDGX agent

arXiv:2607.26339v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) systems ground large language models (LLMs) in external corpora, but this reliance exposes them to corpus poisoning

ReCo: Reweighting GRPO Against Distributional Concentration

Model ReleasesDGX agent

arXiv:2607.26862v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) has become a standard reinforcement learning method for post-training language models. Recent work shows that

Reeling It In: Flexible Needle Pick Up via Thread Manipulation for Autonomous Suturing

Model ReleasesDGX agent

arXiv:2607.26337v1 Announce Type: new Abstract: Suture-needle pickup is necessary for autonomous suturing, as a needle can be unexpectedly dropped or strategically released to adjust the grasping conf

Rethinking Self-Evolution: A Constrained Exploration-Exploitation Process for Mitigating Skill Overfitting

Model ReleasesDGX agent

arXiv:2607.26643v1 Announce Type: cross Abstract: Enabling large language model (LLM) agents to accumulate and reuse experience from past interactions remains a central challenge in real-world applica

Risk-Aware Motion Planning with Learned Trajectory Primitives and Probabilistic Safety Assessment

Model ReleasesDGX agent

arXiv:2607.26802v1 Announce Type: new Abstract: This paper presents a radial basis function network (RBFN)-informed motion planning framework for safe and efficient urban autonomous driving. The propo

Route by Kinematics, Act by Observation: Kinematics-Supervised Expert Routing in MoE-Augmented VLA

Model ReleasesDGX agent

arXiv:2607.26807v1 Announce Type: new Abstract: While MoE augments VLA via expert specialization, router suffers from ineffective expert routing owing to the kinematic heterogeneity of actions across

Same Evidence, Different Target: Decoding How Diagnostic Evidence Bears on Causal Questions from Language-Model States

Model ReleasesDGX agent

arXiv:2607.26929v1 Announce Type: new Abstract: The same diagnostic result can support or challenge one causal claim yet fail to address another when the claims concern different populations, outcomes

SciFigQual-Bench: A Benchmark for Scientific Figure Quality Assessment with Full-Manuscript Context

Model ReleasesDGX agent

arXiv:2607.27084v1 Announce Type: new Abstract: Scientific images are the core elements of presenting experimental conclusions, elaborating system architecture, and supporting comparative arguments in

SecRespond: Benchmarking AI Agents for Real-World Post-Compromise Incident Response

Model ReleasesDGX agent

arXiv:2607.26791v1 Announce Type: cross Abstract: Large Language Model (LLM) agents are increasingly adopted in real-world security operations with access to host artifacts and command-line interfaces

Seeing or Knowing? Visual Context Sensitivity in Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2607.26326v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) achieve strong performance by integrating visual inputs with the rich priors of pretrained language models. How

Semantic-Aware Temporal Adaptation for UAV Anti-UAV Tracking

Model ReleasesDGX agent

arXiv:2607.26511v1 Announce Type: new Abstract: UAV Anti-UAV tracking is an emerging low-altitude security task for localizing an adversarial UAV using the onboard camera of a moving observer UAV. It

SERPO: Self-Evolving Rubric Policy Optimization for Open-Ended Test-Time Reinforcement Learning

Model ReleasesDGX agent

arXiv:2607.26873v1 Announce Type: new Abstract: Test-time reinforcement learning (TTRL) enables language models to self-evolve at inference time without labeled feedback. Existing methods rely on answ

Setoka: A Benchmark for Hierarchical User Understanding in Personalized Agents over Heterogeneous Data

Model ReleasesDGX agent

arXiv:2607.27056v1 Announce Type: cross Abstract: Personalized agents are increasingly applied to assist users across a wide range of tasks. Effective personalized assistance requires not only retriev

Shared SFT Lessons Across Alignment, Model Organisms, and Toy Models

Model ReleasesDGX agent

arXiv:2607.26173v1 Announce Type: new Abstract: Alignment training, model organisms, and toy models are usually treated as separate research areas. But projects in all three frequently use supervised

Simultaneous Coverage and Efficiency Guarantee in Online Conformal Prediction

Model ReleasesDGX agent

arXiv:2607.26577v1 Announce Type: new Abstract: Adaptive conformal inference (ACI) of Gibbs and Cand{es and its variants are the standard approach to online conformal prediction under distribution shi

SpatialQ: Understanding 3D Gaussian Splatting Scene Quality via Visual-based MLLM

Model ReleasesDGX agent

arXiv:2607.26595v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has emerged as an effective representation for novel view synthesis and 3D scene reconstruction, creating an increasing dem

Structurally Separated Uncertainty in Supervised Latent Variable Models

Model ReleasesDGX agent

arXiv:2602.11219v2 Announce Type: replace Abstract: Predictive uncertainty is commonly decomposed into epistemic and aleatoric components, but standard decompositions often produce strongly correlated

sure we lose money on every inference but we make it up in volume

Model ReleasesDGX agent

sure we lose money on every inference but we make it up in volume We are committed to pushing the model frontier across cost efficiency, capability, and speed. Starting today, we are reducing prices f

Symphony of Bias: Exploring Gender Associations with Musical Instruments in Multimodal LLMs

Model ReleasesDGX agent

arXiv:2607.26355v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly embedded in everyday life and widely used for information seeking, raising concerns about their potential

Temporally Centered SIGReg Improves Multi-Task LeWorldModel Learning: From Analysis to Method

Model ReleasesDGX agent

arXiv:2607.26924v1 Announce Type: new Abstract: Recent work on LeWorldModel (LeWM) has shown that the Sketched Isotropic Gaussian Regularizer (SIGReg) enables stable end-to-end world-model learning fr

The Art of Not Forgetting A Local Learning Architecture for Continual Learning

Model ReleasesDGX agent

arXiv:2607.26523v1 Announce Type: new Abstract: We introduce CMP (Cognitive Memory Primitive), a continual-learning architecture that repre?sents inputs as sparse relational codes, stores them in a tw

The Reliability of LLMs for Medical Diagnosis: An Examination of Consistency, Manipulation, and Contextual Awareness

Model ReleasesDGX agent

arXiv:2503.10647v2 Announce Type: replace Abstract: This study evaluated the diagnostic reliability of two Large Language Models (LLMs), Google Gemini 2.0 Flash and OpenAI ChatGPT-4o, across three dim

Thinking Machines just released Inkling-Small: 276B total, 12B active. A faster Inkling that matches or beats its 975B big brother in many b…

Model ReleasesDGX agent

Thinking Machines just released Inkling-Small: 276B total, 12B active. A faster Inkling that matches or beats its 975B big brother in many benchmarks. To test its speed, we plugged it into HF's speech

This is absolutely wild... Anthropic reviewed their logs and found out that their own supposedly-sandboxed cyber evals had hacked three sepa…

Model ReleasesDGX agent

This is absolutely wild... Anthropic reviewed their logs and found out that their own supposedly-sandboxed cyber evals had hacked three separate companies back in April without them noticing! In a rev

Tight Generalization Bound for AdaBoost

Model ReleasesDGX agent

arXiv:2607.26838v1 Announce Type: new Abstract: In this paper we show that the generalization error of AdaBoost is Thetaig(frac{dln(ngamma^{2}/d)}{ngamma^2}+frac{ln(1/elta)}{n}ig), where gamma is the

TiPToP: A Modular Open-Vocabulary Robot Manipulation System That Plans

Model ReleasesDGX agent

arXiv:2603.09971v2 Announce Type: replace Abstract: We present TiPToP, a modular manipulation system that integrates pretrained foundation models with a GPU-accelerated Task and Motion Planner to solv

Towards Trustworthy Embodied Intelligence: A Systems Framework and Graded Trustworthiness Levels

Model ReleasesDGX agent

arXiv:2607.26121v1 Announce Type: new Abstract: Embodied intelligence integrates learned perception and decision making with real-time computation, control, and physical interaction. Because failures

ToxScreen: Detecting Whether an LLM Has Been Poisoned

Model ReleasesDGX agent

arXiv:2607.26849v1 Announce Type: cross Abstract: As large language models (LLMs) are deployed in high-stakes domains, adversaries may poison training data to implant backdoors: hidden triggers that c

TPCD: Tone-Pressure Contrastive Decoding and the Label-Free Gating Bottleneck in Vision-Language Models

Model ReleasesDGX agent

arXiv:2607.26536v1 Announce Type: new Abstract: High-pressure prompts can push vision-language models (VLMs) into unsupported commitments, such as reading illegible text, reporting indeterminate times

Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation

Model ReleasesDGX agent

arXiv:2510.00192v3 Announce Type: replace Abstract: Low-rank adaptation (LoRA) has become a widely used paradigm for parameter-efficient fine-tuning of large language models, yet its representational

Transformers Can Learn Rules They've Never Seen: Proof of Computation Beyond Interpolation

Model ReleasesDGX agent

arXiv:2603.17019v2 Announce Type: replace Abstract: A central question in the debate over large language models is whether transformers can learn rules they have never seen, or whether they can only i

TreeCCA: Canonical Correlation Analysis via Gradient-Boosted Trees

Model ReleasesDGX agent

arXiv:2607.27027v1 Announce Type: new Abstract: Gradient-boosted trees dominate tabular machine learning, yet canonical correlation analysis has always relied on linear or neural encoders. We propose

TREK: A Travel Reasoning and Evaluation Kit for LLM Agents in Complex Trip Planning

Model ReleasesDGX agent

arXiv:2607.26977v1 Announce Type: new Abstract: Travel planning is a demanding stress test for tool-using LLM agents: a usable itinerary is a single artifact that must be right along many axes at once

Turbo-fieldfare: Open-source engine running Gemma 4 26B in 2 GB RAM on Apple Silicon

Model ReleasesDGX agent

Its a custom Swift/Metal inference engine that runs Gemma 4 26B-A4B-IT on M-series Macs with very low RAM. It uses ~2GB instead of ~14 GB. The result is reportedly 5–6 tok/s on an 8 GB M2 MacBook Air

← Previous
1…5152535455…373
Next →