Apple Machine Learning Research

apple.com

Imported from official source

Apple Machine Learning Research — publications from its own official source.

Type
Company
Scope
US · national
Website
apple.com
Feed
Atom

Publications 21

  1. Limits of Confidence in Diffusion

    Discrete diffusion, including remasking and uniform-state samplers, generate a sequence by writing multiple token positions per step, drawing each from a per-position distribution and choosing whic...

    AI
  2. Language Discrimination Improves Linguistic Learning in Multilingual Speech Models

    Multilingual self-supervised speech models can benefit from sharing information across languages, but under a matched total pretraining data budget they still fall short of monolingual models. We s...

    AI
  3. RLTL;DR: Self-Improvement by Internalizing Self-Generated Feedback

    The common paradigm of reinforcement learning with verifiable rewards (RLVR) is to let agents make multiple attempts at a task, and optimize towards the successful ones. This becomes problematic in...

    AI
  4. How Much of a Harness Does a Strong Agent Need for Autonomous ML Engineering?

    Recent autonomous machine learning engineering (MLE) agents have made significant progress on public leaderboards. Often motivated by progress stagnation over long-horizon cycles and limited Large ...

    AI
  5. SCLATE: A Substrate for Continual-Learning Agent Training and Evaluation

    Continual-learning agents are systems of models, harnesses, and memory operating over long multi-session horizons. Evaluating and training them requires interleaving tasks with agent-side events su...

    AI
  6. On the Effectiveness-Fluency Trade-Off in LLM Conditioning: A Systematic Study

    Controlling the output of Large Language Models (LLMs) is a central challenge for their reliable deployment, yet a clear understanding of the involved trade-offs remains elusive. Current approaches...

    AI
  7. The Communication Bottleneck: A Round-Trip Study of Tree-Structured Expression Serialization in Language Models

    When language models reason in chain-of-thought or exchange free-text intermediates, they serialize structured information into natural language. How much tree-structured compositional content surv...

    AI
  8. Faster Rates for Federated Variational Inequalities

    In this paper, we study federated optimization for solving stochastic variational inequalities (VIs), a problem that has attracted growing attention in recent years. Despite substantial progress, a...

    AI
  9. A Practical Recipe for Semi-Supervised Federated ASR: Online Pseudo-Labels with Server Update Stabilization

    Semi-supervised federated learning (SSFL) trains models on clients’ unlabeled data using a teacher to generate pseudo-labels, with a small labeled seed dataset on the server. Automatic Speech Recog...

    AI
  10. Compressing Streaming Neural Audio Encoders via Latent-Space Distillation

    System-wide Dictation on Apple devices runs entirely on-device, and the speech it transcribes reaches the foundation model through a tokenizer: an encoder that maps short windows of waveform onto t...

    AI
  11. How to Guide Your Language Flow

    We introduce a new method to guide flow matching models. Our approach, which we call probe guidance, uses the frozen internal states of an existing diffusion model to construct a guidance signal. T...

    AI
  12. Dynamically Scaled Activation Steering

    Activation steering has emerged as a powerful method for guiding the behavior of generative models towards desired outcomes such as toxicity mitigation. However, most existing methods apply interve...

    AI
  13. REVERSAL-BENCH: A Reversibility Axis and Reset Oracle for Measuring the Reset-Free RL Cliff

    A central goal of autonomous reinforcement learning is continuous policy training without external resets. However, existing paradigms largely depend on underlying environmental reversibility, a pr...

    AI
  14. How Value Induction Reshapes LLM Behaviour

    Conversational Large Language Models are post-trained on language that expresses specific behavioural traits, such as curiosity, open-mindedness, and empathy, and values, such as helpfulness, harml...

    AI
  15. Trajectory as the Teacher: Few-Step Discrete Flow Matching via Energy-Navigated Distillation

    Discrete flow matching generates text by iteratively transforming noise tokens into coherent language, but may require hundreds of forward passes. Distillation uses the multi-step trajectory to tra...

    AI
  16. DACA-GRPO: Denoising-Aware Credit Assignment for Reinforcement Learning in Diffusion Language Models

    Diffusion large language models are a compelling alternative to autoregressive models, yet existing RL methods for diffusion treat all denoising steps as equally important and rely on biased, high-...

    AI
  17. Glyph: A Multi-Strategy Agentic System for Column Description and Sensitivity-Ontology Tagging of Enterprise Data Catalogs

    Enterprise data lakes accumulate tables faster than human stewards can document or classify them, leaving columns with missing descriptions and unassigned governance labels. This documentation debt...

    AI
  18. Shared Selective Persistent Memory for Agentic LLM Systems

    Agentic LLM systems that generate code through multi-turn tool use face a fundamental context problem: each session starts from zero, discarding the configuration choices, domain constraints, data ...

    AI
  19. SimpleDesign: A Joint Model for Protein Sequence and Structure Codesign

    Proteins are fundamental to biological processes, with their function determined by the complex interplay between the amino acid sequence and the three-dimensional structure. Developing generative ...

  20. Putting Captions to the Test: Evaluating Video Caption Quality through Multiple-Choice Question Answering

    Evaluating video captioning remains a critical challenge for Visual Large Language Models (VLLMs). Existing metrics primarily rely on matching generated text against ground-truth references. This p...

    AI
  21. DiscoSign: Discourse-Aware Text to Sign Language Gloss Translation

    Sign language processing systems have traditionally operated at the sentence level, ignoring critical discourse phenomena fundamental to sign language comprehension. We introduce DiscoSign, a compu...

    AI

Officially records where a publication came from, not whether it is true. Imported records are reproduced from an organization's own official source.