AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

OPINE-World: Programmatic World Modeling with Ontology-error-Prioritized Interactive Exploration

DGX agent

arXiv:2607.01531v1 Announce Type: new Abstract: Learning how an environment behaves from interaction is central to building agents that adapt to unfamiliar tasks. World models learned with deep networ

model-releasesarxiv-cs-ai
3 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

OrbitQuant: Data-Agnostic Quantization for Image and Video Diffusion Transformers

DGX agent

arXiv:2607.02461v1 Announce Type: cross Abstract: Diffusion transformers (DiTs) achieve state-of-the-art image and video generation, but their multi-step sampling and growing parameter count make infe

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

PACE: A Proxy for Agentic Capability Evaluation

DGX agent

arXiv:2607.02032v1 Announce Type: new Abstract: Evaluating LLM agents on benchmarks like SWE-Bench and GAIA can be expensive, time-consuming, and requires complex infrastructure. A single evaluation c

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

PairCoder++: Pair Programming as a Universal Paradigm for Verified Code-Driven Multimodal and Structured-Artifact Generation

DGX agent

arXiv:2607.01883v1 Announce Type: new Abstract: Code is the medium through which large language models generate structured artifacts: charts, scientific figures, vector graphics, CAD models, 3D scenes

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

Parameter Golf: What Really Works?

DGX agent

arXiv:2607.01517v1 Announce Type: new Abstract: How far can a language model improve under a strict artifact budget? Parameter Golf posed this question as an open community challenge in which particip

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

Path-level Hindsight Instructions for Semantic Exploration in Vision-Language Navigation

DGX agent

arXiv:2607.01754v1 Announce Type: new Abstract: On-policy exploration is a crucial component for training robust Vision-Language Navigation agents, as it exposes the policy to a broader state distribu

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Phonikud: Overcoming Phonetic Underspecification for Hebrew Text-To-Speech

DGX agent

arXiv:2506.12311v4 Announce Type: replace Abstract: Text-to-speech (TTS) for Modern Hebrew is challenged by the language's orthographic complexity, with existing solutions ignoring underspecified phon

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

PhysMani: Physics-principled 3D World Model for Dynamic Object Manipulation

DGX agent

arXiv:2607.01938v1 Announce Type: cross Abstract: Manipulating fast and dynamically moving targets in unstructured 3D environments remains challenging for embodied AI. Existing visual-language-action

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Population-Scale Segmentation of Penile Tissue in DIXON MRI using Deep Learning for Quantitative Phenotyping in Male Reproductive Health

DGX agent

arXiv:2607.02127v1 Announce Type: cross Abstract: Penile measurement is clinically relevant across male reproductive and urogenital health, including conditions such as micropenis, congenital and endo

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

Power Systems Agent Benchmark: Executable Evaluation of AI Agents in Electric Power Engineering

DGX agent

arXiv:2606.20950v2 Announce Type: replace Abstract: Executable evaluation -- checking the consequences of an agent's actions with a program rather than grading its prose -- has become a prominent way

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

PPTArena: A Benchmark for PowerPoint Editing

DGX agent

arXiv:2512.03042v3 Announce Type: replace-cross Abstract: We introduce PPTArena, a benchmark for PowerPoint editing that evaluates how agents modify real slides from natural-language instructions. Unl

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Pre-Flight: A Benchmark for Evaluating Large Language Models on Aviation Operational Knowledge

DGX agent

arXiv:2607.01829v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly proposed for aviation business operations, from documentation and training generation to customer facing a

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

PreScience: A Dataset and Benchmark for Scientific Forecasting

DGX agent

arXiv:2602.20459v2 Announce Type: replace Abstract: Can AI systems trained on the existing scientific record forecast the advances that will follow? We introduce PreScience, a dataset and benchmark fo

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Program-as-Weights: A Programming Paradigm for Fuzzy Functions

DGX agent

arXiv:2607.02512v1 Announce Type: cross Abstract: Many everyday programming tasks resist clean rule-based implementation, such as alerting on important log lines, repairing malformed JSON, or ranking

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Prompt Framing Distorts Count-Based Evaluation of LLM Error Detection: Evidence from Numeric Anchoring

DGX agent

arXiv:2607.01240v1 Announce Type: cross Abstract: Count-based F1 is widely used as a proxy for LLM error-detection quality, but this paper shows that it can rise dramatically without a corresponding i

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Psychological Imagination Networks Show Cross-Population Centrality and Clustering Alignment in Humans That Large Language Models Fail to Replicate

DGX agent

arXiv:2510.04391v5 Announce Type: replace Abstract: Mental imagery vividness is a stable individual trait, yet whether imagined scenarios share relational structure across human and synthetic large la

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Psychological Steering in LLMs: An Evaluation of Effectiveness and Trustworthiness

DGX agent

arXiv:2510.04484v2 Announce Type: replace-cross Abstract: The ability to control LLMs' emulated emotional states and personality traits is an essential step in enabling rich, human-centered interactio

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

QFedAgent: Quantum-Enhanced Personalized Federated Learning for Multi-Agent Activity Recognition

DGX agent

arXiv:2607.02426v1 Announce Type: cross Abstract: Federated learning (FL) enables collaborative model training across distributed devices without sharing raw data, making it suitable for privacy-sensi

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

RadiomicNet: A Hybrid Radiomics-Guided Lightweight Architecture for Interpretable Medical Image Segmentation

DGX agent

arXiv:2607.02185v1 Announce Type: cross Abstract: Deep learning has achieved remarkable performance in medical image segmentation, yet it suffers from critical limitations: mathematical intractability

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Reasoning LLM Improves Speaker Recognition in Long-form TV Dramas

DGX agent

arXiv:2607.02504v1 Announce Type: cross Abstract: Long-form TV dramas present a formidable challenge for comprehensive video understanding, where deciphering complex storyline often relies on extbf{sp

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Restoring Linguistic Grounding in VLA Models via Train-Free Attention Recalibration

DGX agent

arXiv:2603.06001v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models enable robots to perform manipulation tasks directly from natural language instructions and are increasing

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Robust for the Wrong Reasons: The Representational Geometry of LLM Robustness to Science Skepticism

DGX agent

arXiv:2607.01951v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly consulted on contested scientific questions, raising the concern that they will sycophantically retreat

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

RusFinChain: A Russian Benchmark for Verifiable Chain-of-Thought Reasoning in Finance with Fuzzy-Aligned Evaluation

DGX agent

arXiv:2607.01388v1 Announce Type: new Abstract: Multi-step symbolic reasoning is essential for robust financial analysis, yet most benchmarks neglect intermediate reasoning steps. FINCHAIN introduced

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

SAB-LVLM: Significance-Aware Binarization for Large Vision-Language Models

DGX agent

arXiv:2607.01876v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have achieved remarkable progress in multimodal understanding, yet their enormous parameter scale and cross-modal

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Safety Targeted Embedding Exploit via Refinement

DGX agent

arXiv:2607.01859v1 Announce Type: new Abstract: Safety training for large language models (LLMs) is conducted predominantly in English, leaving uncertain how well safety mechanisms generalize to low-r

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Safety Testing LLM Agents at Scale: From Risk Discovery to Evidence-Grounded Verification

DGX agent

arXiv:2607.01793v1 Announce Type: new Abstract: LLM agents increasingly perform autonomous actions through external tools, leading to complex and evolving safety risks. However, existing safety testin

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Scaling Trends for Lie Detector Oversight in Preference Learning

DGX agent

arXiv:2607.01567v1 Announce Type: new Abstract: Deceptive behavior in LLMs is costly to monitor and prevent, motivating approaches such as Scalable Oversight via Lie Detectors (SOLiD) (Cundy & Gleave,

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Scaling with Confidence: Calibrating Confidence of LLMs for Adaptive Test Time Scaling

DGX agent

arXiv:2607.01612v1 Announce Type: new Abstract: Training large language models (LLMs) with reinforcement learning (RL) has significantly advanced their performance on reasoning and question-answering

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

SCAPE: Accurate and Efficient LLM Training with Extreme Sparse Communication

DGX agent

arXiv:2607.01678v1 Announce Type: new Abstract: Communication increasingly dominates the cost of Large Language Model (LLM) pre-training, especially under data-parallel and sharded training schemes, w

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

Self-Gating Attention for Efficient Time Series Forecasting

DGX agent

arXiv:2607.02344v1 Announce Type: cross Abstract: Transformer architectures have shown strong potential in time series forecasting, where multi-head self-attention is widely used to capture temporal d

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Separating Expert Retention from Autonomous Source Inference in Raw-ECG-Replay-Free Continual ECG Deployment

DGX agent

arXiv:2607.01674v1 Announce Type: new Abstract: In multi-source ECG deployment, models may need to incorporate new data sources when earlier raw ECGs cannot be retained or replayed. Freezing a pretrai

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

SimWorlds: A Multi-Agent System for Dynamic 3D Scene Creation

DGX agent

arXiv:2607.01766v1 Announce Type: new Abstract: LLM agents are increasingly used to translate natural language into 3D scenes in a procedural way, but existing systems focus on static output. Dynamic

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Spec-AUF: Accept-Until-Fail Training under Train-Inference Misalignment for Masked Block Drafters

DGX agent

arXiv:2607.01893v1 Announce Type: new Abstract: Speculative decoding accelerates autoregressive generation by drafting a block of tokens that the target model verifies left-to-right, committing only t

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Spectral Imbalance Causes Forgetting in Low-Rank Continual Adaptation

DGX agent

arXiv:2602.00722v2 Announce Type: replace Abstract: Parameter-efficient continual learning aims to adapt pre-trained models to sequential tasks without forgetting previously acquired knowledge. Most e

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

SPLIT: Cross-Lingual Empathy and Cultural Grounding in English and Ukrainian LLM Responses

DGX agent

arXiv:2607.02049v1 Announce Type: cross Abstract: Large Language Models are increasingly deployed in emotional-support contexts and crisis-related situations. Nevertheless, their cross-lingual abiliti

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

SPOT: Spatio-Temporal Obstacle-free Trajectory Planning for UAVs in Unknown Dynamic Environments

DGX agent

arXiv:2602.01189v3 Announce Type: replace Abstract: We address the problem of reactive motion planning for quadrotors operating in unknown environments with dynamic obstacles. Our approach leverages a

model-releasesarxiv-cs-ro
3 Jul 2026
Model Releases

StatEval: A Comprehensive Benchmark for Large Language Models in Statistics

DGX agent

arXiv:2510.09517v2 Announce Type: replace Abstract: Despite rapid advances in large language models (LLMs), statistical reasoning remains underrepresented in existing LLM benchmarks, which often do no

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

Steerability via constraints: a substrate for scalable oversight of coding agents

DGX agent

arXiv:2607.02389v1 Announce Type: new Abstract: Coding agents are capable; human oversight is the bottleneck. Unconstrained agents introduce security risks, erode codebase scalability, and make human

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Structured Gaussian Processes for Uncertainty-Aware Classification of High-Dimensional, Small-Sampled Omics Data

DGX agent

arXiv:2607.02103v1 Announce Type: cross Abstract: Classifying heterogeneous omics data remains a fundamental challenge in computational biology, particularly in high-dimensional, small-sample settings

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

TestEvo-Bench: An Executable and Live Benchmark for Test and Code Co-Evolution

DGX agent

arXiv:2607.02469v1 Announce Type: cross Abstract: Software tests and code evolve together: a code change should be followed by new or updated tests that record the new software behavior. Yet existing

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Text-Driven 3D Indoor Scene Synthesis in Non-Manhattan Environments

DGX agent

arXiv:2607.02407v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in 3D indoor synthesis for Manhattan environments. However, existing methods ofte

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

The Eticas AI Risk Taxonomy: Open Infrastructure for Operationalizing AI Audits

DGX agent

arXiv:2607.02201v1 Announce Type: cross Abstract: The rapid deployment of AI systems across high-stakes domains has created urgent demand for standardized evaluation, yet the field remains fragmented

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

The Wiola Architecture for Efficient Small Language Models

DGX agent

arXiv:2607.01394v1 Announce Type: new Abstract: We present Wiola, a fully original Small Language Model (SLM) architecture built from first principles, sharing no structural lineage with any existing

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Token Geometry

DGX agent

arXiv:2607.01455v1 Announce Type: cross Abstract: Language models learn continuous programs over discrete symbols, with the embedding table and LM-head acting as the read/write interface between them.

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Towards a Phonology-Informed Evaluation of Multilingual TTS

DGX agent

arXiv:2607.01965v1 Announce Type: new Abstract: Neural TTS systems can sound natural across languages, but naturalness does not guarantee the preservation of sound contrasts that distinguish words fro

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

Towards Load-Aware Prefill Deflection for Disaggregated LLM Serving

DGX agent

arXiv:2607.02043v1 Announce Type: cross Abstract: Disaggregated LLM serving runs prefill and decode on separate GPU pools to keep the two phases from interfering. In practice, this creates a new asymm

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Towards Robustness against Typographic Attack with Training-free Concept Localization

DGX agent

arXiv:2607.02494v1 Announce Type: cross Abstract: Models trained via Contrastive Language-Image Pretraining (CLIP) serve as the foundational vision encoders for most modern Large Vision Language Model

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

TUDUM: A Turkish-Thinking Reasoning Pipeline for Qwen3.5-27B

DGX agent

arXiv:2607.01927v1 Announce Type: cross Abstract: This paper presents TUDUM (Turkce Dusunen Uretken Model), a project pipeline for adapting a Qwen-family 27B thinking model toward Turkish reasoning. T

model-releasesarxiv-cs-ai
3 Jul 2026
← Previous
1…9495969798…361
Next →