AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

87,678Total entries
1Added by human
87,677Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,033 results
7 Jul 2026

PAGE: Towards Practical Human-level Gaze Target Estimation

ResearchDGX agent

arXiv:2607.04860v1 Announce Type: new Abstract: Gaze target estimation, the task of predicting where a person is looking in a scene, is crucial to understanding human attention and intent. It is a cha

ParEVO: Synthesizing Code for Irregular Data: High-Performance Parallelism through Agentic Evolution

Model ReleasesDGX agent

arXiv:2603.02510v2 Announce Type: replace Abstract: The transition from sequential to parallel computing is essential for modern high-performance applications but is hindered by the steep learning cur

PDFBench: A Benchmark for De novo Protein Design from Function

Model ReleasesDGX agent

arXiv:2505.20346v3 Announce Type: replace-cross Abstract: Function-guided protein design is a crucial task with significant applications in drug discovery and enzyme engineering. However, the field la

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

PPE-Bench: A Benchmark for Evaluating MLLM Unlearning under Private-Public Entanglement

Model ReleasesDGX agent

arXiv:2607.02897v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have shown strong capabilities, but they may memorize private information from web data, raising privacy conc

Probing Geospatial SSL Representations with Environmental Signals

Model ReleasesDGX agent

arXiv:2607.05207v1 Announce Type: new Abstract: Self-supervised learning (SSL) is designed to learn generic, transferable representations rather than representations optimized for a single task. Most

Quantize the Target, Quantize the Drafter: Efficient Inference with Qwen3.5-4B

Model ReleasesDGX agent

arXiv:2607.04244v1 Announce Type: new Abstract: This report describes our approach to the Efficient Qwen Competition, where the goal is to enable low-latency serving of Qwen3.5-4B on a resource-constr

Reading Between the Dots: Decoding Hidden Computation across Filler Tokens

Model ReleasesDGX agent

arXiv:2607.03502v1 Announce Type: cross Abstract: Frontier LLMs can perform multi-step reasoning over content-free filler tokens like dots or counting sequences, producing correct answers with no visi

ReLo-IRR: Reflection-Guided LoRA Framework for Image Reflection Removal

Model ReleasesDGX agent

arXiv:2607.02957v1 Announce Type: new Abstract: Single-image reflection removal (SIRR) aims to recover the clean transmission layer from a reflection-contaminated image. Although recent methods achiev

Report: 83% of organizations need to upgrade their infrastructure to support agentic AI

Model ReleasesDGX agent

For years, enterprise AI has been synonymous with conversational AI — the customer service bots and digital assistants we interact with every day. But today, the market has shifted. We’ve officially m

Risk-Constrained Freshness-Aware Semantic Caching for Open-Web Retrieval-Augmented LLMs

Model ReleasesDGX agent

arXiv:2607.04281v1 Announce Type: cross Abstract: Semantic caching reduces the latency and cost of retrieval-augmented generation (RAG) by serving cached answers to semantically similar queries, but m

SABLE: An NDA-Safe Closed-Loop LLM Framework for Analog Circuit Optimization in Industrial EDA Flows

SafetyDGX agent

arXiv:2607.03701v1 Announce Type: cross Abstract: Large language models (LLMs) can propose circuit-optimization decisions, but industrial analog flows cannot expose foundry PDK content, proprietary sc

Serious Games: Human-AI Interaction, Evolution, and Coevolution

ResearchDGX agent

arXiv:2505.16388v2 Announce Type: replace Abstract: The serious games between humans and AI have only just begun. Evolutionary Game Theory (EGT) models the competitive and cooperative strategies of bi

Shapley-based Data Valuation for LLM Alignment via Sequential Preference Optimization

SafetyDGX agent

arXiv:2512.15765v3 Announce Type: replace Abstract: Data valuation is a natural framework for understanding which preference datasets matter most when aligning a Large Language Model (LLM) using multi

Shifting from Discrete to Continuous Reference Data: QSM-Derived Horizontal Tree Biomass Distribution for Deep Learning Biomass Estimation

ResearchDGX agent

arXiv:2607.05260v1 Announce Type: cross Abstract: Conventional modeling approaches for LiDAR-based above-ground biomass (AGB) estimation rely on discrete plot-level inventory aggregates. This methodol

Shortcut Learning in Legal Judgment Prediction: Empirical Evidence from the UK Employment Tribunal

ApplicationsDGX agent

arXiv:2607.04261v1 Announce Type: new Abstract: Current Legal Judgment Prediction (LJP) is constrained by its reliance on post-hoc judicial materials, increasing the likelihood that models perform ret

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses

SafetyDGX agent

arXiv:2510.15476v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly used as interfaces to information, code, and real-world services, making prompt-level security f

sqlite-utils 4.0, now with database schema migrations

Model ReleasesDGX agent

This morning I released sqlite-utils 4.0, the 124th release of that project and the first major version bump since 3.0 in November 2020. In addition to some small but significant breaking changes (des

The Changing Role of Symbolic Methods in Artificial Intelligence

AgentsDGX agent

arXiv:2607.05168v1 Announce Type: new Abstract: Why do intelligent systems need to perform explicit symbolic reasoning? Computer science has traditionally regarded symbolic reasoning as a defining com

Toward Efficient Agents: Memory, Tool learning, and Planning

Model ReleasesDGX agent

arXiv:2601.14192v2 Announce Type: replace Abstract: Recent years have witnessed increasing interest in extending large language models into agentic systems. While the effectiveness of agents has conti

Unsupervised Features Mining via Activation Geometry

ResearchDGX agent

arXiv:2607.04222v1 Announce Type: new Abstract: Interpretability methods aim to reveal the features represented inside large language models (LLMs). Many existing methods begin with labeled examples o

6 Jul 2026

128 GB of memory is nice, but you can get started with local agentic AI workflows with much less. By connecting gemma 4 in @lmstudio to MATL…

Model ReleasesDGX agent

128 GB of memory is nice, but you can get started with local agentic AI workflows with much less. By connecting gemma 4 in @lmstudio to MATLAB MCP Server, you can run a local AI model that uses MATLAB

During a Bloomberg interview, Yann LeCun (@ylecun ) explains why LLMs are limited in terms of real-world intelligence during a Bloomberg int…

Model ReleasesDGX agent

During a Bloomberg interview, Yann LeCun (@ylecun ) explains why LLMs are limited in terms of real-world intelligence during a Bloomberg interview. 'Language is a very approximate, reduced, quantized,

Interesting stuff. And the visualization at the end is worth trying: https://www.neuronpedia.org/qwen3.6-27b/jlens

Model ReleasesDGX agent

Interesting stuff. And the visualization at the end is worth trying: https://www.neuronpedia.org/qwen3.6-27b/jlens New Anthropic research: A global workspace in language models. Of everything happenin

Segmental Attention Decoding with Long Form Acoustic Encodings

TutorialsDGX agent

We address the fundamental incompatibility of attention-based encoder-decoder (AED) models with long-form acoustic encodings. AED models trained on segmented utterances learn to encode absolute frame

5 Jul 2026

Check out all the amazing work from our @SimonsFdn Collaboration on the Physics of Learning and Neural Computation (https://www.physicsoflea…

TutorialsDGX agent

Check out all the amazing work from our @SimonsFdn Collaboration on the Physics of Learning and Neural Computation (https://www.physicsoflearning.org/) presented at the main meeting of @ICMLconf #ICML

3 Jul 2026

AbsoluteDegradation: A Physics-Inspired Synthetic Film-Degradation Pipeline and Archival Film Restoration Benchmark

Model ReleasesDGX agent

arXiv:2607.02131v1 Announce Type: cross Abstract: Restoring archival film remains a fundamentally challenging problem due to the absence of paired training data and the lack of standardized evaluation

Activation Steering for Aligned Open-ended Generation without Sacrificing Coherence

Model ReleasesDGX agent

arXiv:2604.08169v2 Announce Type: replace Abstract: Alignment in LLMs is more brittle than commonly assumed: misalignment can be induced by adversarial prompts, benign fine-tuning, emergent misalignme

Composite Reward Design in PPO-Driven Adaptive Filtering

Model ReleasesDGX agent

arXiv:2506.06323v2 Announce Type: replace-cross Abstract: Model-free and reinforcement learning-based adaptive filtering methods are gaining traction for denoising in dynamic, non-stationary environme

ContextSniper: AntTrail's Token-Efficient Code Memory for Repository-Level Program Repair

Model ReleasesDGX agent

arXiv:2607.01916v1 Announce Type: new Abstract: Large language model agents can repair real repository issues, but they often spend large context budgets on whole-file reads, broad searches, and long

Denser neq Better: Limits of On-Policy Self-Distillation for Continual Post-Training

Model ReleasesDGX agent

arXiv:2607.01763v1 Announce Type: cross Abstract: Continual post-training enables foundation models to acquire new knowledge while preserving existing capabilities. Recent work suggests that on-policy

Do Newer Lightweight CNNs Perform Better Under Resource Constraints? A Controlled Multigenerational Study of Architecture, Initialization, Training Budget, and Efficiency

Model ReleasesDGX agent

arXiv:2607.01984v1 Announce Type: cross Abstract: Newer lightweight convolutional neural networks are often presented as improving predictive performance and deployment efficiency, but such claims req

EPnG: Adaptive Expert Prune-and-Grow for Parameter-Efficient MoE Fine-tuning

Model ReleasesDGX agent

arXiv:2607.01789v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models scale efficiently but remain costly to adapt due to redundant experts and uniform parameter allocation. Existing param

Excited to share our paper, “Learning Multi-Agent Coordination via Sheaf-ADMM” to be presented at #ICML2026 Blog: https://pub.sakana.ai/shea…

Model ReleasesDGX agent

Excited to share our paper, “Learning Multi-Agent Coordination via Sheaf-ADMM” to be presented at #ICML2026 Blog: https://pub.sakana.ai/sheaf-admm/ Most AI models process information as one giant, mon

Introduction to Transformers: an NLP Perspective

ResearchDGX agent

arXiv:2311.17633v2 Announce Type: replace-cross Abstract: Transformers have dominated empirical machine learning models of natural language processing. In this paper, we introduce basic concepts of Tr

LACUNA: A Testbed for Evaluating Localization Precision for LLM Unlearning

Model ReleasesDGX agent

arXiv:2607.02513v1 Announce Type: cross Abstract: LLMs memorize sensitive training data, including personally identifiable information (PII), creating a pressing need for reliable post hoc removal met

MMIR-TCM: Memory-Integrated Multimodal Inference and Retrieval for TCM Clinical Decision Support

Model ReleasesDGX agent

arXiv:2607.01814v1 Announce Type: new Abstract: Traditional Chinese Medicine (TCM) diagnosis, particularly through tongue inspection, faces persistent challenges in subjectivity and reproducibility. T

One More Time: Revisiting Neural Quantum States from a Reinforcement Learning Perspective

Model ReleasesDGX agent

arXiv:2607.02292v1 Announce Type: new Abstract: Neural quantum states (NQS) provide a flexible and scalable framework for approximating quantum many-body wavefunctions. Among NQS parameterizations, au

Overthink-Triggered Slowdown Attacks on LVLM-Based Robotic Systems

SafetyDGX agent

arXiv:2607.01518v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have been increasingly integrated into robotic systems. However, these models may exhibit overthinking behaviors,

PACE: A Neuro-Symbolic Framework for Plausible and Actionable Counterfactual Explanations

ApplicationsDGX agent

arXiv:2607.01306v1 Announce Type: new Abstract: Counterfactual explanations explain machine learning predictions by identifying minimal input changes that would alter a model's decision. Although many

Phonikud: Overcoming Phonetic Underspecification for Hebrew Text-To-Speech

Model ReleasesDGX agent

arXiv:2506.12311v4 Announce Type: replace Abstract: Text-to-speech (TTS) for Modern Hebrew is challenged by the language's orthographic complexity, with existing solutions ignoring underspecified phon

Population-Scale Segmentation of Penile Tissue in DIXON MRI using Deep Learning for Quantitative Phenotyping in Male Reproductive Health

Model ReleasesDGX agent

arXiv:2607.02127v1 Announce Type: cross Abstract: Penile measurement is clinically relevant across male reproductive and urogenital health, including conditions such as micropenis, congenital and endo

Scaling Trends for Lie Detector Oversight in Preference Learning

Model ReleasesDGX agent

arXiv:2607.01567v1 Announce Type: new Abstract: Deceptive behavior in LLMs is costly to monitor and prevent, motivating approaches such as Scalable Oversight via Lie Detectors (SOLiD) (Cundy & Gleave,

Self-explainable Operator Learning for Discovering Spatial Patterns in Functional Data

Local AiDGX agent

arXiv:2607.02203v1 Announce Type: new Abstract: Operator learning has emerged as a powerful tool for modeling complex physical systems in functional spaces. However, their neural network-based archite

Separating Expert Retention from Autonomous Source Inference in Raw-ECG-Replay-Free Continual ECG Deployment

Model ReleasesDGX agent

arXiv:2607.01674v1 Announce Type: new Abstract: In multi-source ECG deployment, models may need to incorporate new data sources when earlier raw ECGs cannot be retained or replayed. Freezing a pretrai

SPARCLE: SPeaker-aware Aligned Representations via Contrastive Language Embeddings

ResearchDGX agent

arXiv:2607.01238v1 Announce Type: cross Abstract: Recent advances in speech synthesis have shifted from phoneme representations to direct grapheme modeling. While phonemes address the one-to-many mapp

Spec-AUF: Accept-Until-Fail Training under Train-Inference Misalignment for Masked Block Drafters

Model ReleasesDGX agent

arXiv:2607.01893v1 Announce Type: new Abstract: Speculative decoding accelerates autoregressive generation by drafting a block of tokens that the target model verifies left-to-right, committing only t

SPLIT: Cross-Lingual Empathy and Cultural Grounding in English and Ukrainian LLM Responses

Model ReleasesDGX agent

arXiv:2607.02049v1 Announce Type: cross Abstract: Large Language Models are increasingly deployed in emotional-support contexts and crisis-related situations. Nevertheless, their cross-lingual abiliti

Token Geometry

Model ReleasesDGX agent

arXiv:2607.01455v1 Announce Type: cross Abstract: Language models learn continuous programs over discrete symbols, with the embedding table and LM-head acting as the read/write interface between them.

UA-ChatDev: Uncertainty-Aware Multi-Agent Collaboration for Reliable Software Development

Model ReleasesDGX agent

arXiv:2607.02186v1 Announce Type: new Abstract: Software development is a complex task that demands cooperation among agents with diverse roles. Large language models (LLMs) have enabled autonomous mu

2 Jul 2026

A Filtered Mixture-of-Generators for Fully Synthetic Survival Training

SafetyDGX agent

arXiv:2607.00127v1 Announce Type: new Abstract: Survival analysis models time-to-event data, but in clinical settings training data are costly and scarce: events accrue over years of follow-up, cohort

AGE: Adaptive-masking for Graph Embedding in Graph Retrieval-Augmented Generation

Model ReleasesDGX agent

arXiv:2607.00052v1 Announce Type: cross Abstract: GraphRAG is an extension of retrieval-augmented generation (RAG) that supports large language models (LLMs) by referring to graph-structured data as e

An LLM-Based Framework for Intent-Driven Network Topology Design

Model ReleasesDGX agent

arXiv:2607.00292v1 Announce Type: cross Abstract: Designing deployable and resilient network topologies from natural language requirements remains a challenging problem in network automation. This wor

Beyond Perplexity: A Behavioral Evaluation Framework for Deployment-Memory Claims in LLM Test-Time Training

Local AiDGX agent

arXiv:2607.00368v1 Announce Type: new Abstract: Large language model test-time training (TTT) is often evaluated through local proxy metrics: models are updated on recent tokens, retrieved context, ta

Can Agents Generalize to the Open World? Unveiling the Fragility of Static Training in Tool Use

Model ReleasesDGX agent

arXiv:2607.01084v1 Announce Type: new Abstract: While Large Language Model (LLM) agents demonstrate proficiency in static benchmarks, their deployment in real-world scenarios is hindered by the dynami

Continual learning is probably the biggest barrier to explosive AI adoption (& may have big implications for recursive self-improvement as w…

Model ReleasesDGX agent

Continual learning is probably the biggest barrier to explosive AI adoption (& may have big implications for recursive self-improvement as well) As long as you deal with amnesiac models that require h

Does Your ViT Still Need U-Net for Segmentation?

Model ReleasesDGX agent

arXiv:2607.00223v1 Announce Type: new Abstract: Medical image segmentation is dominated by U-Net-style encoder-decoder architectures. Vision Transformers (ViTs) overcome the limited receptive field of

Flow-Map GRPO: Reinforcement Learning for Few-Step Flow-Map Generators via Anchored Stochastic Composition

ResearchDGX agent

arXiv:2607.00535v1 Announce Type: cross Abstract: Few-step flow-map generators, such as consistency models and MeanFlow, accelerate sampling by directly learning long-range transport maps between nois

FLYNN: Robust Neural Network for Robot Navigation using Fly Brain Topology

Model ReleasesDGX agent

arXiv:2607.00025v1 Announce Type: cross Abstract: While deep learning models achieve state-of-the-art performance in complex tasks, they remain brittle when faced with new environments or sensory depr

GLM 5.2 DSpark preview is here! ✨ https://huggingface.co/RedHatAI/GLM-5.2-speculator.dspark-preview This is the first DSpark speculator for …

Model ReleasesDGX agent

GLM 5.2 DSpark preview is here! ✨ https://huggingface.co/RedHatAI/GLM-5.2-speculator.dspark-preview This is the first DSpark speculator for a non-DeepSeek frontier model, trained with Speculators and

GMO-E^2DIT: Grounded Multi-Operation Editing for E-Commerce Images

Model ReleasesDGX agent

arXiv:2607.00920v1 Announce Type: new Abstract: Real-world e-commerce image editing often requires multiple, localized, and auditable operations rather than global restyling. This compositional nature

← Previous
1…389390391392393…1051
Next →