AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
All
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

Search: “apple-ml-research”

GridTimelineEvolution
49+ results
Research

Arbitrage: Efficient Reasoning via Advantage-Aware Speculation

DGX agent

Modern Large Language Models achieve impressive reasoning capabilities with long Chain of Thoughts, but they incur substantial computational cost during inference, and this motivates techniques to imp

researchapple-ml-research
7 Aug 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

Beyond Next-Token Prediction: A Performance Characterization of Diffusion versus Autoregressive Language Models

DGX agent

Large Language Models (LLMs) have achieved state-of-the-art performance on a broad range of Natural Language Processing (NLP) tasks, including document processing and code generation. Autoregressive L

researchapple-ml-research
7 Aug 2026
Research

Scaling Categorical Flow Maps

DGX agent

Continuous diffusion and flow matching models could represent a powerful alternative to autoregressive approaches for language modelling (LM), as they unlock a host of advantages currently reserved fo

researchapple-ml-research
7 Aug 2026
Research

DeepAmbigQA: Ambiguous Multi-hop Questions for Benchmarking LLM Answer Completeness

DGX agent

Large language models (LLMs) with integrated search tools show strong promise in open-domain question answering (QA), yet they often struggle to produce complete answer set to complex questions such a

researchapple-ml-research
6 Aug 2026
Research

Locking Pretrained Weights via Deep Low-Rank Residual Distillation

DGX agent

The quality of open-weight language models has dramatically improved in recent years. Sharing weights greatly facilitates model adoption by enabling their use across diverse hardware and software plat

researchapple-ml-research
6 Aug 2026
Safety

Understanding Alignment in Multimodal LLMs: A Comprehensive Study

DGX agent

Preference alignment has become a crucial component in enhancing the performance of Large Language Models (LLMs), yet its impact in Multimodal Large Language Models (MLLMs) remains comparatively under

safetyapple-ml-research
3 Aug 2026
Local Ai

Memory Efficient Audio Synthesis with Decoupled Temporal Depth Diffusion Transformers

DGX agent

Siri Expressive Voices synthesize rich, configurable speech in real time and entirely on device, powered by AFM 3 Core Advanced, Apple’s most powerful on-device foundation model. This work presents th

local-aiapple-ml-research
28 Jul 2026
Research

GH-ESD: Grounded Hypothesis-Driven Error Slice Discovery for Instance-Level Vision Tasks

DGX agent

Systematic failures of vision models on semantically coherent subsets, known as error slices, reveal limitations in robustness and evaluation. Existing slice discovery approaches largely model slices

researchapple-ml-research
27 Jul 2026
Research

LEAD: Breaking the No-Recovery Bottleneck in Long-Horizon Reasoning

DGX agent

Long-horizon execution in Large Language Models (LLMs) remains unstable even when high-level strategies are provided. Evaluating on controlled algorithmic puzzles, we demonstrate that while decomposit

researchapple-ml-research
24 Jul 2026
Research

Accelerating Text-to-Video Generation with Calibrated Sparse Attention

DGX agent

Recent diffusion models enable high-quality video generation, but suffer from slow runtimes. The large transformer-based backbones used in these models are bottlenecked by spatiotemporal attention. In

researchapple-ml-research
21 Jul 2026
Research

CLaRa: Bridging Retrieval and Generation with Continuous Latent Reasoning

DGX agent

Retrieval-augmented generation (RAG) enhances large language models (LLMs) with external knowledge but still suffers from long contexts and disjoint retrieval–generation optimization. In this work, we

researchapple-ml-research
15 Jul 2026
Research

One Layer Is Enough: Adapting Pretrained Visual Encoders for Image Generation

DGX agent

Visual generative models (e.g., diffusion models) typically operate in compressed latent spaces to balance training efficiency and sample quality. In parallel, there has been growing interest in lever

researchapple-ml-research
15 Jul 2026
Model Releases

Multilingual Semantic Retrieval for Apple Music Search

DGX agent

Apple Music serves listeners across 150+ storefronts in dozens of languages, with a catalog that grows by hundreds of thousands of new tracks daily. At this scale, search recall on misspelled, transli

model-releasesapple-ml-research
14 Jul 2026
Agents

Proactive Agent Research Environment: Simulating Active Users to Evaluate Proactive Assistants

DGX agent

Proactive agents that anticipate user needs and autonomously execute tasks hold great promise as digital assistants, yet the lack of realistic user simulation frameworks hinders their development. Exi

agentsapple-ml-research
14 Jul 2026
Agents

Behavioral Privacy Leakage in Agentic Negotiation: Formalizing and Mitigating Inference Attacks via Randomized Policies

DGX agent

This paper was accepted at the AI4TCI (Workshop on AI for Secure and Trustworthy Critical Infrastructure Systems) Workshop at the International Conference on Availability, Reliability and Security (AR

agentsapple-ml-research
10 Jul 2026
Safety

Incentivizing Temporal-Awareness in Egocentric Video Understanding Models

DGX agent

Multimodal large language models (MLLMs) have recently shown strong performance in visual understanding, yet they often lack temporal awareness, particularly in egocentric settings where reasoning dep

safetyapple-ml-research
9 Jul 2026
Agents

Recursive Language Models Meet Uncertainty: The Surprising Effectiveness of Self-Reflective Program Search for Long Context

DGX agent

Long-context handling remains a core challenge for language models: even with extended context windows, models often fail to reliably extract, reason over, and use the information across long contexts

agentsapple-ml-research
9 Jul 2026
Safety

Unmasking On-Policy Distillation: Where It Helps, Where It Hurts, and Why

DGX agent

On-policy distillation offers dense, per-token supervision for training reasoning models; however, it remains unclear under which conditions this signal is beneficial and under which it is detrimental

safetyapple-ml-research
9 Jul 2026
Research

FlowEval: Reference-Based Evaluation of Generated User Interfaces

DGX agent

While large language models (LLMs) and coding agents are often applied to user interface (UI) development, developers find it difficult to reliably assess their proficiency in visual and interaction d

researchapple-ml-research
7 Jul 2026
Safety

MT-EditFlow: Reinforcement Learning for Multi-Turn Image Editing with Flow Matching

DGX agent

Recent breakthroughs in instruction-based image editing have captured significant attention, as models are now capable of handling real-world editing demands with the practicality required by everyday

safetyapple-ml-research
7 Jul 2026
Research

Taming Text-to-Sounding Video Generation via Advanced Modality Condition and Interaction

DGX agent

This study focuses on Text-to-Sounding-Video (T2SV) generation, which aims to generate a video with synchronized audio from text, with both modalities aligned to the text conditions. Despite progress

researchapple-ml-research
7 Jul 2026
Research

Weblica: Scalable and Reproducible Training Environments for Visual Web Agents

DGX agent

The web is complex, open-ended, and constantly changing, making it challenging to scale training data for visual web agents. Existing data collection attempts remain limited to offline trajectories fo

researchapple-ml-research
7 Jul 2026
Research

Path-Constrained Mixture-of-Experts

DGX agent

Sparse Mixture-of-Experts (MoE) architectures route each token through a subset of experts at each layer independently. We propose viewing MoE computation through the lens of expert paths—the sequence

researchapple-ml-research
6 Jul 2026
Research

Revisiting ASR Error Correction with Specialized Models

DGX agent

Language models play a central role in automatic speech recognition (ASR), yet most methods rely on text-only models unaware of ASR error patterns. Recently, large language models (LLMs) have been app

researchapple-ml-research
6 Jul 2026
Tutorials

Segmental Attention Decoding with Long Form Acoustic Encodings

DGX agent

We address the fundamental incompatibility of attention-based encoder-decoder (AED) models with long-form acoustic encodings. AED models trained on segmented utterances learn to encode absolute frame

tutorialsapple-ml-research
6 Jul 2026
Safety

Understanding Annotator Safety Policy with Interpretability

DGX agent

Safety policies define what constitutes safe and unsafe AI outputs, guiding data annotation and model development. However, annotation disagreement is pervasive and can stem from multiple sources such

safetyapple-ml-research
6 Jul 2026
Research

Amortizing Maximum Inner Product Search with Learned Support Functions

DGX agent

Maximum inner product search (MIPS) is a crucial subroutine in machine learning, requiring the identification of a vector taken within a database (the keys) that best aligns with a given query. We pro

researchapple-ml-research
2 Jul 2026
Research

Anti-Causal Domain Generalization: Leveraging Unlabeled Data

DGX agent

The problem of domain generalization concerns learning predictive models that are robust to distribution shifts when deployed in new, previously unseen environments. Existing methods typically require

researchapple-ml-research
2 Jul 2026
Research

Learning Structured Reasoning via Tractable Trajectory Control

DGX agent

Large language models can exhibit emergent reasoning behaviors, often manifested as recurring lexical patterns (e.g., “wait,” indicating verification). However, complex reasoning trajectories remain s

researchapple-ml-research
2 Jul 2026
Research

Learning Unmasking Policies for Diffusion Language Models

DGX agent

Diffusion (Large) Language Models (dLLMs) now match the downstream performance of their autoregressive counterparts on many tasks, while holding the promise of being more efficient during inference. O

researchapple-ml-research
2 Jul 2026
Research

MemoryLLM: Plug-n-Play Interpretable Feed-Forward Memory for Transformers

DGX agent

Understanding how transformer components operate in LLMs is important, as it is at the core of recent technological advances in artificial intelligence. In this work, we revisit the challenges associa

researchapple-ml-research
2 Jul 2026
Agents

Multi-Agent Teams Hold Experts Back

DGX agent

Multi-agent LLM systems are increasingly deployed as autonomous collaborators, where agents interact freely rather than execute fixed, pre-specified workflows. In such settings, effective coordination

agentsapple-ml-research
2 Jul 2026
Research

On Robustness and Chain-of-Thought Consistency of RL-Finetuned VLMs

DGX agent

Reinforcement learning (RL) finetuning has become a key technique for enhancing large language models (LLMs) on reasoning-intensive tasks, motivating its extension to vision language models (VLMs). Wh

researchapple-ml-research
2 Jul 2026
Research

Residual Context Diffusion Language Models

DGX agent

Diffusion Large Language Models (dLLMs) have emerged as a promising alternative to purely autoregressive language models because they can decode multiple tokens in parallel. However, state-of-the-art

researchapple-ml-research
2 Jul 2026
Research

Metric-Dependent Annotation Saturation for Learning from Label Distributions

DGX agent

When annotators disagree on a label, the disagreement itself carries signal—and the number of annotators needed to capture it depends on the evaluation metric. We fine-tune NLI models on label distrib

researchapple-ml-research
23 Jun 2026
Local Ai

Introducing the Third Generation of Apple’s Foundation Models

DGX agent

Our next generation of Apple Intelligence is centered around our users, integrated deeply into our operating systems, and powered by a bold new architecture with privacy at its core. At the heart of t

local-aiapple-ml-research
8 Jun 2026
Research

IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2026

DGX agent

CVPR 2026 is a major international conference on computer vision and pattern recognition organized by IEEE and the Computer Vision Foundation, where Apple presents its latest machine learning research

researchapple-ml-research
28 May 2026
Research

VSAS-Bench: Real-Time Evaluation of Visual Streaming Assistant Models

DGX agent

Streaming vision-language models (VLMs) continuously generate responses given an instruction prompt and an online stream of input frames. This is a core mechanism for real-time visual assistants. Exis

researchapple-ml-research
22 May 2026
Research

Apple Workshop on Privacy-Preserving Machine Learning & AI 2026

DGX agent

At Apple, we believe privacy is a fundamental human right. As AI capabilities increase and become more integrated into people’s daily lives, advancing research in privacy-preserving techniques is incr

researchapple-ml-research
8 May 2026
Safety

RVPO: Risk-Sensitive Alignment via Variance Regularization

DGX agent

Current critic-less RLHF methods aggregate multi-objective rewards via an arithmetic mean, leaving them vulnerable to constraint neglect: high-magnitude success in one objective can numerically offset

safetyapple-ml-research
8 May 2026
Research

Velox: Learning Representations of 4D Geometry and Appearance

DGX agent

We introduce a framework for learning latent representations of 4D objects which are descriptive, faithfully capturing object geometry and appearance; compressive, aiding in downstream efficiency; and

researchapple-ml-research
8 May 2026
Research

What Matters in Practical Learned Image Compression

DGX agent

One of the major differentiators unlocked by learned codecs relative to their hard-coded traditional counterparts is their ability to be optimized directly to appeal to the human visual system. Despit

researchapple-ml-research
7 May 2026
Model Releases

From Where Things Are to What They’re For: Benchmarking Spatial–Functional Intelligence for Multimodal LLMs

DGX agent

True spatial intelligence for multimodal agents transcends low-level geometric perception, evolving from knowing where things are to understanding what they are for. While existing benchmarks, such as

model-releasesapple-ml-research
6 May 2026
Research

SpecMD: A Comprehensive Study on Speculative Expert Prefetching

DGX agent

Mixture-of-Experts (MoE) models enable sparse expert activation, meaning that only a subset of the model’s parameters is used during each inference. However, to translate this sparsity into practical

researchapple-ml-research
6 May 2026
Research

Bootstrapping Sign Language Annotations with Sign Language Models

DGX agent

AI-driven sign language interpretation is limited by a lack of high-quality annotated data. New datasets including ASL STEM Wiki and FLEURS-ASL contain professional interpreters and 100s of hours of d

researchapple-ml-research
30 Apr 2026
Research

International Conference on Acoustics, Speech and Signal Processing (ICASSP) 2026

DGX agent

ICASSP 2026 is a major international conference focused on acoustics, speech, and signal processing research and applications. Apple's machine learning research team is participating in the conference

researchapple-ml-research
30 Apr 2026
Research

STARFlow-V: End-to-End Video Generative Modeling with Normalizing Flows

DGX agent

Normalizing flows (NFs) are end-to-end likelihood-based generative models for continuous data, and have recently regained attention with encouraging progress on image generation. Yet in the video gene

researchapple-ml-research
30 Apr 2026
Research

Adaptive Thinking: Large Language Models Know When to Think in Latent Space

DGX agent

Recent advances in large language models (LLMs) test-time computing have introduced the capability to perform intermediate chain-of-thought (CoT) reasoning (thinking) before generating answers. While

researchapple-ml-research
29 Apr 2026
← Previous
12
Next →
62 results