This is a hashtag index for recurring themes across Reading, Research, and Work.

Hashtags

Click a hashtag to see associated items across sections.

  • #ai-safety - alignment and safety behavior in real model use.
  • #evaluation - what measurements are actually faithful.
  • #model-as-judge - evaluator models and judgment reliability tradeoffs.
  • #rlaif - reinforcement learning from AI feedback and its behavioral effects.
  • #emergent-misalignment - narrow training, broad behavioral drift.
  • #tool-use - tool interfaces, harnesses, and runtime constraints.
  • #agents - long-horizon behavior, oversight, and reliability loops.
  • #multi-turn - reliability and degradation across multi-turn interaction.
  • #decoding - inference-time transformation of next-token distributions.
  • #sampling - token selection strategies and uncertainty control.
  • #heuristics - practical decoding rules used as baselines or supervisors.
  • #ml-systems - training and inference system design.
  • #medical-ai - model evaluation and behavior in clinical contexts.
  • #research-engineering - execution quality from question to evidence.
  • #generalization - behavior under pressure and distribution shift.
  • #benchmarks - benchmark design, scope, and limitations.
  • #programmable-biology - biological design and validation loops.
  • #bioinformatics - computational biology methods and tooling.
  • #synthetic-biology - engineering approaches in biology.
  • #biofoundries - infrastructure for biological build-test cycles.
  • #systems - systems-level thinking across software and research workflows.
  • #faithfulness - plausible outputs versus faithful reasoning.