AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

87,171Total entries
1Added by human
87,170Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
87,171 results
3 Jul 2026

Meta-Benchmarks for Financial-Services LLM Evaluation

Model ReleasesDGX agent

arXiv:2607.01740v1 Announce Type: new Abstract: Public LLM leaderboards optimise for global average performance and do not capture the specific cognitive demands of financial-services work: a model th

Meta could use its compute for its own models, ad scaling, SpaceX-like neocloud deals, and hosting 3rd-party models; it may be close to an Anthropic deal (Jeremie Eliahou Ontiveros/SemiAnalysis)

IndustryDGX agent

Jeremie Eliahou Ontiveros / SemiAnalysis: Meta could use its compute for its own models, ad scaling, SpaceX-like neocloud deals, and hosting 3rd-party models; it may be close to an Anthropic deal — Zu

Meta getting into the cloud business has been inevitable for a long time, as it seeks to diversify beyond ad revenue and monetize its AI buildout (M.G. Siegler/Spyglass)

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Industry
DGX agent

M.G. Siegler / Spyglass: Meta getting into the cloud business has been inevitable for a long time, as it seeks to diversify beyond ad revenue and monetize its AI buildout — Their need to diversify the

Meta-Representational Predictive Coding: Neuroscience-Informed Self-Supervised Learning

ResearchDGX agent

arXiv:2503.21796v2 Announce Type: replace-cross Abstract: Self-supervised learning has become an increasingly important paradigm in the domain of machine intelligence. Furthermore, evidence for self-s

Meta to release new AI model with advanced coding capabilities ‘soon’

Model ReleasesDGX agent

Meta Platforms Inc. is gearing up to release a new version of its flagship Muse Spark artificial intelligence model. Alexandr Wang, the company’s chief AI officer, wrote on X today that the update wil

MetaTT: A Global Tensor-Train Adapter for Parameter-Efficient Fine-Tuning

Model ReleasesDGX agent

arXiv:2506.09105v3 Announce Type: replace-cross Abstract: We present MetaTT, a Tensor Train (TT) adapter framework for fine-tuning of pre-trained transformers. MetaTT enables flexible and parameter-ef

MetaTune: Adjoint-based Meta-tuning via Robotic Differentiable Dynamics

SafetyDGX agent

arXiv:2603.27313v2 Announce Type: replace Abstract: Disturbance observer-based control has shown promise in robustifying robotic systems against uncertainties. However, tuning such systems remains cha

Midjourney wants Disney, Universal, and Warner Bros. to reveal in court how they use AI across their companies; studios sued Midjourney in 2025 for infringement (Gene Maddaus/Variety)

IndustryDGX agent

Gene Maddaus / Variety: Midjourney wants Disney, Universal, and Warner Bros. to reveal in court how they use AI across their companies; studios sued Midjourney in 2025 for infringement — The studios s

Mirror Illusion Art

SafetyDGX agent

arXiv:2607.02015v1 Announce Type: cross Abstract: Mirror Illusion Art is a novel reflection-conditioned 3D illusion where one object yields two target appearances (front and mirror). The task is formu

Mixture-of-Parallelisms: Towards Memory-Efficient Training Stack for Mixture-of-Experts Models

Model ReleasesDGX agent

arXiv:2607.01844v1 Announce Type: cross Abstract: This paper showcases a memory-efficient training stack for Mixture-of-Experts (MoE) models. It is a training paradigm that combines and specializes va

MKGR: Multimodal Knowledge-Graph Representation Learning for Cold-Start Protein-Protein Interaction Prediction

Model ReleasesDGX agent

arXiv:2607.01627v1 Announce Type: cross Abstract: Accurate protein-protein interaction (PPI) prediction is central to functional genomics, disease mechanism discovery, and drug development. A difficul

MMAO-Cls: Metabolic Multi-Agent Optimization for Joint Feature Selection and Classifier Tuning

AgentsDGX agent

arXiv:2607.01539v1 Announce Type: cross Abstract: This paper studies whether the Metabolic Multi-Agent Optimizer (MMAO) can act as a credible outer-loop optimizer for classification model selection. W

MMBench-Live: A Continuously Evolving Benchmark for Multimodal Models

Model ReleasesDGX agent

arXiv:2607.01813v1 Announce Type: cross Abstract: Evaluation benchmarks are essential for assessing vision-language models (VLMs), but most multimodal benchmarks are static, making them vulnerable to

MMIR-TCM: Memory-Integrated Multimodal Inference and Retrieval for TCM Clinical Decision Support

Model ReleasesDGX agent

arXiv:2607.01814v1 Announce Type: new Abstract: Traditional Chinese Medicine (TCM) diagnosis, particularly through tongue inspection, faces persistent challenges in subjectivity and reproducibility. T

Model Merging as Probabilistic Inference in Fine-Tuning Parameter Space

Model ReleasesDGX agent

arXiv:2607.01689v1 Announce Type: cross Abstract: Model merging aims to combine existing single-task solutions into a multi-task solution without additional data-driven fine-tuning.~Most existing appr

MolSight: A Graph-Aware Vision-Language Model for Unified Chemical Image Understanding

SafetyDGX agent

arXiv:2607.01982v1 Announce Type: cross Abstract: Using molecular large language models (LLMs) as a unified framework for understanding molecular structures and functions is emerging as a new trend in

More notes on my blog: https://simonwillison.net/2026/Jul/3/judgement/

ToolsDGX agent

Simon Willison shares additional notes and commentary on his blog regarding judgment and decision-making processes. The post likely explores practical considerations or philosophical perspectives on h

Morphology-Aware Sample Assignment: Overcoming IoU Insensitivity for Surface Defect Detection

SafetyDGX agent

arXiv:2606.13723v2 Announce Type: replace-cross Abstract: Intersection-over-Union (IoU), as a pivotal metric for evaluating the spatial alignment between candidate proposals and ground-truth annotatio

Motion-Focused Latent Action Enables Cross-Embodiment VLA Training from Human EgoVideos

ApplicationsDGX agent

arXiv:2606.18955v2 Announce Type: replace-cross Abstract: Training generalist Vision-Language-Action(VLA) models typically requires massive, diverse robotic datasets with high-fidelity action annotati

MultAttnAttrib: Training-Free Multimodal Attribution in Long Document Question Answering

Model ReleasesDGX agent

arXiv:2607.01420v1 Announce Type: cross Abstract: As grounded QA systems are increasingly deployed in AI assistants, accurately attributing generated answers to evidence is critical for user trust and

Multi-Head Recurrent Memory Agents

ResearchDGX agent

arXiv:2607.01523v1 Announce Type: cross Abstract: Recurrent memory agents extend LLMs to arbitrarily long contexts by iteratively consolidating input into a fixed-size memory window. Despite their sca

Multi-modal Rail Crossing Safety Analysis

SafetyDGX agent

arXiv:2607.01365v1 Announce Type: cross Abstract: Given one or more images of a railway crossing, can we leverage visual cues that allow us to robustly estimate how safe it is? Can we improve our abil

Multi-Objective Exploration and Preference Optimization via Mutual Information

SafetyDGX agent

arXiv:2607.01392v1 Announce Type: new Abstract: Aligning large language models with diverse and heterogeneous human values requires multi-objective alignment methods to effectively trade off conflicti

Multi-Rate Nonlinear Model Predictive Control for Wall-Supported Bipedal Locomotion of Quadrupedal Robots

ResearchDGX agent

arXiv:2607.01574v1 Announce Type: new Abstract: This paper presents a novel layered planning and control framework based on multi-rate nonlinear model predictive control (MR-NMPC) that enables quadrup

Multilayer Q-Matrix-Embedded Neural Network for Cognitive Diagnosis (M-QCDNet): Structure-Aware Deep Learning Architecture for Psychometric Interpretability

SafetyDGX agent

arXiv:2607.01278v1 Announce Type: new Abstract: The research proposes a multilayer Q-matrix-embedded neural network for cognitive diagnosis (M-QCDNet), which integrates the structural interpretability

Multilingual Prompt Localization for Agent-as-a-Judge: Language and Backbone Sensitivity in Requirement-Level Evaluation

Model ReleasesDGX agent

arXiv:2604.04532v2 Announce Type: replace-cross Abstract: Evaluation language is typically treated as a fixed English default in agentic code benchmarks, yet we show that changing the judge's language

Multimodal Knowledge Edit-Scoped Generalization for Online Recursive MLLM Editing

Local AiDGX agent

arXiv:2607.01978v1 Announce Type: new Abstract: Online multimodal knowledge editing requires injecting a continual stream of visual-textual corrections into multimodal large language models (MLLMs) wi

Multimodal prompting is clearly the future. I love experimenting with new ways to interact with agents. As a researcher and engineer, I've f…

AgentsDGX agent

Multimodal prompting is clearly the future. I love experimenting with new ways to interact with agents. As a researcher and engineer, I've found that the richer the inputs to the agent and the richer

mupscaling small models: Principled warm starts and hyperparameter transfer

Model ReleasesDGX agent

arXiv:2602.10545v2 Announce Type: replace-cross Abstract: Modern large-scale neural networks are often trained and released in multiple sizes to accommodate diverse inference budgets. To improve effic

'My belief is that, three years down the road, people will not own PCs. You will stop buying PCs and laptops. You will just have a phone, or…

AgentsDGX agent

'My belief is that, three years down the road, people will not own PCs. You will stop buying PCs and laptops. You will just have a phone, or maybe you might have a tablet.' -- AGI Inc.'s Div Garg on w

NarrativeTrack: Evaluating Entity-Centric Reasoning for Narrative Understanding

Model ReleasesDGX agent

arXiv:2601.01095v3 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) have achieved impressive progress in vision-language reasoning, yet their ability to understand tempo

NAVER LABS Europe Submission to the Instruction-following 2026 Short Track

ResearchDGX agent

arXiv:2607.01960v1 Announce Type: new Abstract: In this paper, we describe NAVER LABS Europe's submission to the instruction-following speech processing short track at IWSLT 2026. We participate again

Navigating the Alignment-Calibration Trade-off: A Pareto-Superior Frontier via Model Merging

SafetyDGX agent

arXiv:2510.17426v3 Announce Type: replace-cross Abstract: The 'alignment tax' of post-training is typically framed as a drop in task accuracy. We show it also involves a severe loss of calibration, ma

NeoMap: Training-free Novel-View Synthesis from Single Images and Videos

SafetyDGX agent

arXiv:2607.01962v1 Announce Type: cross Abstract: We study the challenging problem of novel view video synthesis from single images or monocular videos. Existing methods, which operate under the assum

Neuro-Symbolic Safety Guidance for Vision-Language-Action Models via Constrained Flow Matching

Model ReleasesDGX agent

arXiv:2607.01378v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have demonstrated promising generalization capabilities across robotic manipulation tasks, yet their real-world depl

NeuroBridge: Bridging Multi-Task MRI Knowledge for Neurodegenerative Disease Diagnosis

ResearchDGX agent

arXiv:2607.01401v1 Announce Type: cross Abstract: INTRODUCTION: Accurate MRI-based identification of Alzheimer's disease (AD), mild cognitive impairment (MCI), and related dementias remains challengin

Neuron-Aware Active Few-Shot Learning for LLMs

ResearchDGX agent

arXiv:2607.02423v1 Announce Type: cross Abstract: Active Few-Shot Learning (AFSL) adapts LLMs to specialized domains by identifying the most valuable unlabeled samples for annotation and use as few-sh

Neuron-Aware Data Selection for Annotation-Free LLM Self-Distillation

SafetyDGX agent

arXiv:2607.02460v1 Announce Type: cross Abstract: Post-training large language models (LLMs) without real-world interaction feedback or human-labeled supervision remains challenging, particularly in s

NEUROSYMLAND: Neuro-Symbolic Landing-Site Assessment for Robust and Edge-Deployable UAV Autonomy

Model ReleasesDGX agent

arXiv:2607.02277v1 Announce Type: new Abstract: Safe landing-site assessment in unstructured environments remains a key challenge for autonomous UAV deployment, as vision-only learning approaches ofte

NEW paper worth reading. (bookmark it) The basic idea is to pair a compressive recurrent state with a small exact memory, which helps to rec…

TutorialsDGX agent

NEW paper worth reading. (bookmark it) The basic idea is to pair a compressive recurrent state with a small exact memory, which helps to recover long-range recall without giving up the efficiency of l

Non-synchronism in Global Usage of Research Methods in Library and Information Science from 1990 to 2019

TutorialsDGX agent

arXiv:2607.01833v1 Announce Type: cross Abstract: The global development of Library and Information Science (LIS) is influenced by various factors such as the economy, society, culture, discipline, tr

None of it was an accident. A team that gave up recharge week, and a partnership @GeoffBibby built with @swyx + the AI Engineer crew that tu…

ToolsDGX agent

None of it was an accident. A team that gave up recharge week, and a partnership @GeoffBibby built with @swyx + the AI Engineer crew that turned a year-old idea into the biggest stage in AI eng. Thank

Object Aligner: A Configurable JSON Schema Similarity Score for Graphs, Applied to LLM Prompt Optimization

SafetyDGX agent

arXiv:2607.01972v1 Announce Type: cross Abstract: Large language models (LLMs) are often asked to produce JSON conforming to a fixed schema, powering information extraction, tool calling, agentic plan

Object-centric LeJEPA

ResearchDGX agent

arXiv:2607.02404v1 Announce Type: cross Abstract: Image encoders trained with LeJEPA can deliver strong features for downstream tasks, but, like other image-level self-supervised methods, typically re

Office Comprehension Benchmark

Model ReleasesDGX agent

arXiv:2607.01245v1 Announce Type: cross Abstract: We introduce Office Comprehension Bench (OCB), the first public benchmark to jointly evaluate LLM systems on Word, Excel, and PowerPoint comprehension

OmniGAIA: Towards Native Omni-Modal AI Agents

Model ReleasesDGX agent

arXiv:2602.22897v3 Announce Type: replace Abstract: Human intelligence naturally intertwines omni-modal perception -- spanning vision, audio, and language -- with complex reasoning and tool usage to i

On the Asymptotics of Self-Supervised Pre-training: Two-Stage M-Estimation and Representation Symmetry

TutorialsDGX agent

arXiv:2603.27631v2 Announce Type: replace Abstract: Self-supervised pre-training, where large corpora of unlabeled data are used to learn representations for downstream fine-tuning, has become a corne

On the Dimension-Free Approximation of Deep Neural Networks for Symmetric Korobov Functions

ResearchDGX agent

arXiv:2511.12398v2 Announce Type: replace Abstract: Deep neural networks have been widely used as universal approximators for functions with inherent physical structures, including permutation symmetr

On the Limits of Steering Vectors for Preference-Aligned Generation

Model ReleasesDGX agent

arXiv:2607.01802v1 Announce Type: new Abstract: Steering vectors have emerged as a promising approach to controlled text generation, offering interpretable, training-free mechanisms for shaping model

On the Role of Computation in Reinforcement Learning

SafetyDGX agent

arXiv:2602.05999v4 Announce Type: replace Abstract: How does the amount of compute available to a reinforcement learning (RL) policy affect its learning? Can policies using a fixed amount of parameter

On the Role of Directionality in Structural Generalization

ResearchDGX agent

arXiv:2607.02307v1 Announce Type: new Abstract: Several SLOG test categories explicitly involve directional distinctions (modifier position shifts, argument extraction positions), yet AM-Parser, the p

On the Sample Efficiency of Inverse Dynamics Models for Semi-Supervised Imitation Learning

SafetyDGX agent

arXiv:2602.02762v2 Announce Type: replace Abstract: Semi-supervised imitation learning (SSIL) consists in learning a policy from a small dataset of action-labeled trajectories and a much larger datase

On the Utility and Factual Reliability of Pruned Mixture-of-Experts Models in the Biomedical Domain

Model ReleasesDGX agent

arXiv:2607.01444v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models offer inference speedups via selective activation but impose substantial memory requirements because the whole network

One Demonstration Is Enough for Real-World Robotic Reinforcement Learning

SafetyDGX agent

arXiv:2607.01651v1 Announce Type: new Abstract: Learning effective robot control policies on physical hardware is challenging due to costly data collection and the difficulty of reward specification.

One More Time: Revisiting Neural Quantum States from a Reinforcement Learning Perspective

Model ReleasesDGX agent

arXiv:2607.02292v1 Announce Type: new Abstract: Neural quantum states (NQS) provide a flexible and scalable framework for approximating quantum many-body wavefunctions. Among NQS parameterizations, au

One way to live 400 years is to do (and experience) 5x more things per day

ResearchDGX agent

This post explores a philosophical perspective on subjective time perception, suggesting that by engaging in more diverse activities and novel experiences daily, one can psychologically expand their l

Online Resource Allocation with Continuous Random Consumption: Regret under Degeneracy

SafetyDGX agent

arXiv:2607.02196v1 Announce Type: new Abstract: We study online resource allocation when both rewards and consumption sizes may be continuously distributed. Requests arrive sequentially and must be ac

Online Safety Monitoring for LLMs

SafetyDGX agent

arXiv:2607.02510v1 Announce Type: new Abstract: Despite alignment training, LLMs remain prone to generating unsafe outputs at deployment time. Monitoring outputs online and raising an alarm when safet

OntoLearner: A Modular Python Library for Ontology Learning with Large Language Models

ResearchDGX agent

arXiv:2607.01977v1 Announce Type: new Abstract: Ontology learning (OL) aims to automatically construct structured knowledge models from text, yet progress remains fragmented across methods, domains, a

Open model usage has gone from 10% of AI tokens to 30% in a year. The shift to open, modular AI is here to stay. Our founders on what's driv…

ToolsDGX agent

Open source AI model usage has tripled from 10% to 30% of total AI tokens consumed within a year, reflecting a significant market shift toward open and modular AI architectures. This trend indicates g

← Previous
1…378379380381382…1453
Next →