AivexaNewsSearch
AI news for builders and product teamsUpdated Oct 10, 2026, 23:01 UTC

Lilian Weng

Reporting and perspectives from Lilian Weng. Headlines and excerpts link to the original articles.

Latest stories

Newest first

Controllable Neural Text Generation

Lilian Weng's blog post surveys methods for controlling neural text generation with pretrained language models, covering decoding strategies, guided and trainable decoding, prompt design, and fine-tuning approaches. It was updated through September 2021 to add P-tuning, Prompt Tuning, and unlikelihood training.

Neural Architecture Search

A survey-style report describes Neural Architecture Search (NAS), the automatic design of neural network architectures. It outlines the three components defined by Elsken et al. 2019—search space, search algorithm, and evaluation strategy—and reviews search spaces, search algorithms, and evaluation strategies from prior work.

The Transformer Family

Lilian Weng updated her post on the Transformer family on 2023-01-27, refactoring it to incorporate new Transformer models since 2020. The revised version, titled The Transformer Family Version 2.0, is the recommended reference on the topic.

Curriculum for Reinforcement Learning

A technical overview of curriculum learning for reinforcement learning, covering task-specific curricula, teacher-guided methods, self-play, automatic goal generation, skill-based curricula, and curriculum through distillation. The article notes updates made on 2020-02-03 and 2020-02-04 and cites work including Elman (1993) and Bengio et al. (2009).

Read original

Evolution Strategies

Lilian Weng published a technical post on Evolution Strategies, covering Gaussian ES, CMA-ES parameter updates, Natural Evolution Strategies, and applications in deep reinforcement learning including OpenAI ES and exploration methods.

Meta Reinforcement Learning

Lilian Weng's post surveys Meta Reinforcement Learning, tracing its origin to a 2001 Hochreiter et al. paper and its 2016 proposals by Wang et al. and Duan et al. It covers formulations, key components, and algorithms including meta-gradient RL, Evolved Policy Gradient, MAESN, and episodic control.

Read original