AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,566 results
30 Jun 2026

Core dump epidemiology: fixing an 18-year-old bug

Model ReleasesDGX agent

This article describes OpenAI's investigation and resolution of a longstanding bug in their data infrastructure that had persisted for 18 years, likely related to how core dumps (system memory snapsho

CORTEX: High-Quality Cross-Domain Organization of Web-Scale Corpora through Ontological Corpus Graph

Model ReleasesDGX agent

arXiv:2606.30175v1 Announce Type: new Abstract: The continuous evolution of large language models drives escalating demands on data scale and quality, and as different training stages impose increasin

CoSPlan: Corrective Sequential Planning via Scene Graph Incremental Updates

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2512.10342v3 Announce Type: replace Abstract: Vision Language Models (VLMs) have shown promising planning capabilities, yet their success remains confined to the text domain, leaving visual deci

CostBench: Evaluating Multi-Turn Cost-Optimal Planning and Adaptation in Dynamic Environments for LLM Tool-Use Agents

Model ReleasesDGX agent

arXiv:2511.02734v3 Announce Type: replace Abstract: Current evaluations of Large Language Model (LLM) agents primarily emphasize task completion, often overlooking resource efficiency and adaptability

Counterfactual Residual Data Augmentation for Regression

Model ReleasesDGX agent

arXiv:2606.28460v1 Announce Type: cross Abstract: Data-driven modeling in real-world regression tasks often suffers from limited training samples, high collection costs, and noisy observations. Inspir

Cross-Modal Iteration Distillation for Robust IHD Screening: The IDNet Framework and A New Benchmark

Model ReleasesDGX agent

arXiv:2606.30027v1 Announce Type: new Abstract: Color Fundus Photography (CFP) offers a low-cost and non-invasive route for ischemic heart disease (IHD) screening, but current studies are limited by s

Cross-Temporal Sinhala OCR: Page-Level Adaptation and Diachronic Analysis

Model ReleasesDGX agent

arXiv:2606.29378v1 Announce Type: new Abstract: Sinhala is a morphologically rich abugida spoken by roughly 16 million people in Sri Lanka, and to date, there are no publicly available real-world data

CubifyGS: Object-Centric 3D Gaussian Splatting for Lifelong Dynamic Scene Maintenance

Model ReleasesDGX agent

arXiv:2606.28720v1 Announce Type: new Abstract: Lifelong scene mapping under rigid object rearrangement remains a fundamental challenge in robotics. While 3D Gaussian Splatting (3DGS) enables high-fid

Curvature-Weighted Gradient Diversity: A Noise Measure for Geometry-Adaptive SGD Schedules

Model ReleasesDGX agent

arXiv:2606.30455v1 Announce Type: new Abstract: The standard convergence analysis of mini-batch stochastic gradient descent (SGD) models gradient noise using a single variance term that treats all par

Customized Generative AI Agent for Transportation Engineering Practice: A Development and Continued Pre-training Guideline

Model ReleasesDGX agent

arXiv:2606.29014v1 Announce Type: new Abstract: Recent advancements in generative artificial intelligence (AI) and large language models (LLMs) have shown significant promise in automating complex rea

CylindTrack: Depth-Aware Cylindrical Motion Modeling for Panoramic Multi-Object Tracking

Model ReleasesDGX agent

arXiv:2606.30097v1 Announce Type: new Abstract: Multi-Object Tracking (MOT) is a core capability for embodied perception, and panoramic cameras are attractive for embodied systems because their 360{eg

Data and Evaluation Closed-Loop for Model Capability Enhancement

Model ReleasesDGX agent

arXiv:2606.28471v1 Announce Type: new Abstract: Model capability is the central variable in LLM pre-training, yet is never observed directly: data shapes it prospectively, while evaluation reveals it

Data-Driven Energy-Based Learning via Gibbs Measures on Hierarchical Structures

Model ReleasesDGX agent

arXiv:2606.30064v1 Announce Type: new Abstract: We introduce a data-driven probabilistic framework for learning systems based on Gibbs measures on hierarchical structures. Unlike standard empirical ri

Database Context Compression for Text-to-SQL on Real-World Large Databases

Model ReleasesDGX agent

arXiv:2606.28601v1 Announce Type: cross Abstract: Recent progress in Text-to-SQL has been driven by stronger language models and prompting strategies, yet performance on real enterprise benchmarks suc

DataComp-VLM: Improved Open Datasets for Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.28551v1 Announce Type: cross Abstract: Building performant Vision-Language Models (VLMs) requires carefully curating large-scale training datasets, yet the community lacks systematic benchm

Defeat Devices in AI Systems

Model ReleasesDGX agent

arXiv:2606.28863v1 Announce Type: cross Abstract: AI systems increasingly exhibit behavior that differs systematically between evaluation and deployment contexts. Alignment faking, sandbagging, benchm

Demonstration-Free Robotic Control via LLM Agents

Model ReleasesDGX agent

arXiv:2601.20334v2 Announce Type: replace-cross Abstract: Robotic manipulation has increasingly adopted vision-language-action (VLA) models, which achieve strong performance but typically require task

Detecting Clinical Hallucinations in LVLMs via Counterfactual Visual Grounding Uncertainty

Model ReleasesDGX agent

arXiv:2606.28520v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) are increasingly used for clinical image understanding, yet they remain vulnerable to hallucinations--producing t

DIA-HARM: Dialectal Disparities in Harmful Content Detection Across 50 English Dialects

Model ReleasesDGX agent

arXiv:2604.05318v2 Announce Type: replace Abstract: Harmful content detectors, particularly disinformation classifiers, are predominantly developed and evaluated on Standard American English (SAE), le

Diagnosing and Mitigating Retrieval Bottlenecks in LLM-Based Cold-Start Recommendation

Model ReleasesDGX agent

arXiv:2606.29947v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as rerankers in recommender systems, with the expectation that semantic understanding will help in

DialogPII: A multilingual dataset of synthetic dialog transcripts to detect personal information

Model ReleasesDGX agent

arXiv:2606.30312v1 Announce Type: new Abstract: Conversational data collected in domains such as healthcare or social sciences is a valuable resource for research and automated analysis. However, resp

Diff-Based Code Corruption using LLMs for Large-Scale Bugfix Benchmarking

Model ReleasesDGX agent

arXiv:2606.29088v1 Announce Type: cross Abstract: There are various benchmarks to evaluate bugfixing capabilities of Large Language Models. However, most widespread benchmarks do not fully reflect rea

Digitizing Coaching Intelligence: An Agentic Framework for Holistic Athlete Profiling using VLM and RAG

Model ReleasesDGX agent

arXiv:2606.28570v1 Announce Type: cross Abstract: Athlete assessment is a critical process for tracking physical progress and identifying elite talent. However, during mass recruitment drives, traditi

DiscoGen: Procedural Generation of Algorithm Discovery Tasks in Machine Learning

Model ReleasesDGX agent

arXiv:2603.17863v2 Announce Type: replace-cross Abstract: Automating the development of machine learning algorithms has the potential to unlock new breakthroughs. However, our ability to improve and e

DistilledGemma: Balanced Efficiency-Accuracy for Person-Place Relation Extraction from Multilingual Historical Articles

Model ReleasesDGX agent

arXiv:2606.29130v1 Announce Type: new Abstract: We present DistilledGemma, an efficient and accurate system for the HIPE-2026 shared task on person-place relation extraction from multilingual historic

Diversity is the Strength of the AI Crowd

Model ReleasesDGX agent

arXiv:2606.29661v1 Announce Type: new Abstract: Top AI forecasting systems are approaching superforecaster-level accuracy on future world events, but still rely primarily on off-the-shelf LLMs combine

DLR: Zero-Inference-Cost Latent Residuals for Low-Rank Pre-Training

Model ReleasesDGX agent

arXiv:2606.28932v1 Announce Type: cross Abstract: Large language models have driven recent progress in language and multimodal AI, yet pre-training them at scale is prohibitively expensive. Low-rank p

DNA Language Models: An Assessment of Pre-Training for Fine-Tuning Tasks

Model ReleasesDGX agent

arXiv:2606.30140v1 Announce Type: cross Abstract: Recent breakthroughs in foundation models and Large Language Models (LLMs) have introduced new opportunities for studying and decoding genomic sequenc

Dockerless: Environment-Free Program Verifier for Coding Agents

Model ReleasesDGX agent

arXiv:2606.28436v1 Announce Type: cross Abstract: Program verifiers play a central role in training coding agents, including selecting trajectories for supervised fine-tuning (SFT) and providing rewar

Does Verbose Chain-of-Thought Really Help? In-Distribution Evidence that Content, Not Length, Matters

Model ReleasesDGX agent

arXiv:2606.30128v1 Announce Type: new Abstract: Chain-of-thought (CoT) prompting improves LLM reasoning, but the source is contested: do the intermediate steps help because they carry useful semantic

Domain-Informed Multi-View Self-Distillation for Astronomical Light-Curve Representation Learning with JEPA

Model ReleasesDGX agent

arXiv:2606.28446v1 Announce Type: cross Abstract: Light curves describe temporal variations in the brightness of celestial objects. Learning robust representations of light curves is essential for lar

DR-GS: Physically-Based Deformable and Relightable 2D Gaussians

Model ReleasesDGX agent

arXiv:2606.29379v1 Announce Type: cross Abstract: Gaussian splatting (GS) has garnered significant attention in VR/AR and digital content creation due to its explicit parameterization and efficient re

DuoMem: Towards Capable On-Device Memory Agents via Dual-Space Distillation

Model ReleasesDGX agent

arXiv:2606.29961v1 Announce Type: cross Abstract: Large Language Model (LLM)-based agents can solve complex procedural tasks by interacting with environments over multiple turns, but this ability typi

Dynamic Parsing and Updating Natural Language Specification using VLMs for Robust Vision-Language Tracking

Model ReleasesDGX agent

arXiv:2606.29357v1 Announce Type: cross Abstract: Vision-language tracking guided by natural language specifications leverages high-level semantic cues of target objects to substantially boost trackin

Dynamo: Dynamic Skill-Tool Evolution for Vision-Language Agents

Model ReleasesDGX agent

arXiv:2606.30185v1 Announce Type: new Abstract: Improving vision-language models (VLMs) on visual reasoning typically requires retraining or hand-designed prompts and tools. We present Dynamo, a train

Earlier this month, the City of Starbase was excited to support residents, local fishermen, @SpaceX volunteers, and @SeaTurtleInc in rescuin…

Model ReleasesDGX agent

Earlier this month, the City of Starbase was excited to support residents, local fishermen, @SpaceX volunteers, and @SeaTurtleInc in rescuing 266-lb loggerhead Arlen from Boca Chica Beach jetties. She

Early Cue Precision Shapes Visual Shortcut Learning in Controlled Cue-Manipulation Benchmarks

Model ReleasesDGX agent

arXiv:2606.30344v1 Announce Type: cross Abstract: Visual classifiers can achieve high matched-distribution accuracy while relying on low-level cues that fail under conflict or suppression. We test whe

Early Estimation of Language to Latent Alignment in Diffusion Models

Model ReleasesDGX agent

arXiv:2512.08505v2 Announce Type: replace Abstract: Conditional diffusion models frequently suffer from language-image misalignments. Due to the ambiguity of intermediate noise corrupted latents, asse

Early Warning Signals for OpenVLA Failure under Visual Distribution Shift

Model ReleasesDGX agent

arXiv:2606.29699v1 Announce Type: cross Abstract: Vision Language Action models combine perception, language grounding, and control in a single policy, but their failures are hard to diagnose once vis

Echoes of Human Malice in Agents: Benchmarking LLMs for Multi-Turn Online Harassment Attacks

Model ReleasesDGX agent

arXiv:2510.14207v3 Announce Type: replace Abstract: Large Language Model (LLM) agents are powering a growing share of interactive web applications, yet remain vulnerable to misuse and harm. Prior jail

Efficient Visual Pointing for Embodied AI:Agent-Driven Data Synthesis, Cross-Block Attention, and Iterative Correction

Model ReleasesDGX agent

arXiv:2606.29850v1 Announce Type: new Abstract: Visual pointing maps a language instruction to pixel co ordinates, a core skill for embodied AI. We describe our PointArena 2026 solution, which achieve

Emergence of Minimal Circuits for Indirect Object Identification in Attention-Only Transformers

Model ReleasesDGX agent

arXiv:2510.25013v2 Announce Type: replace-cross Abstract: Mechanistic interpretability aims to reverse-engineer large language models (LLMs) into human-understandable computational circuits. However,

EMPATH: A Multilingual Auditor-Judge Benchmark for Safety Evaluation of Emotional-Support Chatbots

Model ReleasesDGX agent

arXiv:2606.30256v1 Announce Type: new Abstract: Safety benchmarks often buy scalability by fixing the prompt, the language, and the turn structure. For emotional-support chatbots, that bargain hides p

Ensemble Learning Based Classification Algorithm Recommendation

Model ReleasesDGX agent

arXiv:2101.05993v2 Announce Type: replace-cross Abstract: Selecting an appropriate classification algorithm for a given data set remains a challenging problem in data mining and machine learning. Exis

Entropy Regularized Reinforcement Learning for Zero-Sum Stochastic Differential Games in a Regime-Switching Jump-Diffusion Process

Model ReleasesDGX agent

arXiv:2606.28669v1 Announce Type: new Abstract: To address parameter misspecification and sudden structural environmental changes in conventional stochastic differential game (SDG) frameworks, this pa

Envisage: Diffusion-Based Rhinoplasty Goal Visualization with Mask-Decomposed Evaluation

Model ReleasesDGX agent

arXiv:2606.28628v1 Announce Type: cross Abstract: Localized generative editing needs localized evaluation: full-image identity metrics are structurally confounded under hard-composited edits. We prese

EVAF: A Test-Retest Protocol for Selective Parametric Consolidation

Model ReleasesDGX agent

arXiv:2606.29916v1 Announce Type: cross Abstract: Long-running language agents need mechanisms for deciding which experiences should persist after the working context is gone. Retrieval systems can re

Eval-Actions: Fine-Grained Execution Quality Evaluation for Robotic Manipulation

Model ReleasesDGX agent

arXiv:2601.18723v2 Announce Type: replace Abstract: Although Vision--Action (VA) and Vision--Language--Action (VLA) policies have advanced robotic manipulation, their evaluation remains dominated by b

EvalSafetyGap: A Hybrid Survey and Conceptual Framework for LLM Evaluation-Safety Failures

Model ReleasesDGX agent

arXiv:2606.30219v1 Announce Type: new Abstract: LLM evaluation and AI safety face a shared measurement problem: benchmark scores, reward-model signals, and reported safety metrics can improve while th

Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions

Model ReleasesDGX agent

arXiv:2507.05257v4 Announce Type: replace-cross Abstract: Recent benchmarks for Large Language Model (LLM) agents primarily focus on evaluating reasoning, planning, and execution capabilities, while a

EventVLA: Event-Driven Visual Evidence Memory for Long-Horizon Vision-Language-Action Policies

Model ReleasesDGX agent

arXiv:2606.20092v2 Announce Type: replace Abstract: Memory remains a critical bottleneck for long-horizon robotic manipulation, as standard Vision-Language-Action (VLA) policies often fail when task-r

Every article saying 'THIS IS THE NEW JOB OF THE AI ERA' is one of two jobs. And the first is a lie imo. The first job, is usually something…

Model ReleasesDGX agent

Every article saying 'THIS IS THE NEW JOB OF THE AI ERA' is one of two jobs. And the first is a lie imo. The first job, is usually something that an AI lab is hiring for and the news blows it complete

EVLA: An Electro-Aware Multimodal Assistant for Physically-Grounded Driving Reasoning and Control

Model ReleasesDGX agent

arXiv:2606.28938v1 Announce Type: new Abstract: Modern vision-language models (VLMs) for driving assistants typically treat vehicle dynamics as a black box, resulting in decisions that lack awareness

Evolutional Math: Cross-Validated Island-Model Genetic Programming for Interpretable Symbolic Regression on Small, Wide Datasets

Model ReleasesDGX agent

arXiv:2606.28381v1 Announce Type: cross Abstract: Symbolic regression via genetic programming routinely fails on small, wide datasets - a regime common in clinical-trial monitoring, biostatistics, and

Experience Augmented Policy Optimization for LLM Reasoning

Model ReleasesDGX agent

arXiv:2606.30420v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is a powerful paradigm for improving the reasoning capabilities of large language models (LLMs). H

Expert Evaluation of Clinical AI Tools on Real Point-of-Care Clinical Queries

Model ReleasesDGX agent

arXiv:2606.28960v1 Announce Type: new Abstract: Physicians now pose millions of clinical questions to AI tools each week, yet these tools are evaluated largely on hypothetical or exam-style questions,

Explainability-Aware Frustum Attack: Exposing Structural Vulnerabilities in LiDAR-Based 3D Object Detectors

Model ReleasesDGX agent

arXiv:2606.29963v1 Announce Type: new Abstract: The structural vulnerabilities of point cloud-based 3D object detectors remain poorly understood. Prior work has studied adversarial robustness primaril

Explaining Attention with Program Synthesis

Model ReleasesDGX agent

arXiv:2606.19317v2 Announce Type: replace-cross Abstract: A longstanding goal of research on interpretable deep learning is to replace opaque neural computations with human-meaningful symbolic descrip

Exploiting Local Flatness for Efficient Out-of-Distribution Detection

Model ReleasesDGX agent

arXiv:2606.29952v1 Announce Type: cross Abstract: Detecting out-of-distribution (OOD) data is crucial for reliable machine learning deployment. Among detection strategies, post-hoc methods are particu

Explore ideas, scale visual concepts, and start creating: http://goo.gle/4bcThNt

Model ReleasesDGX agent

This Google AI post likely promotes a tool or platform for exploring creative ideas, scaling visual concepts, and beginning creative projects, with a shortened URL linking to more details. The post ap

← Previous
1…117118119120121…377
Next →