principles.fyi · The Brain

The Brain

Every concept in the library, wired by the articles that share them. It's dark when you arrive — reading lights it up. Click any node.

a line = two concepts in the same article · node size = how many articles use it

Every concept, A–Z

anisotropyattentionautoregressionbackpropagationbase modelBayes' ruleBERTbidirectionalBIO taggingBradley-Terry modelcausal maskchain of thoughtcompute-optimal scalingcloze task[CLS] tokenCommon Crawlconditional generationcontext windowcontextual embeddingcosine similaritycross-entropy lossdata contaminationdecoderdenoisingderivativedot productDPOeembeddingencoderencoder-decoderentropyexpected valueexponentialfew-shotfeed-forward networkfine-tuningfoundation modelGeLUgenerative AIgrouped-query attentiongradient descentgreedy decodinghallucinationhidden statein-context learninginduction headinstruction tuningKL divergenceKV cachelayer normlearning ratelogarithmlog-oddslogit lenslogitsLoRA[MASK] tokenmasked language modelingmatrix multiplymeanmin-p samplingmixture of expertsMMLUmodel alignmentmulti-head attentionnamed entity recognitionnatural language inferencenext sentence predictionnonlinearityoddsparametersrepetition / presence / frequency penaltyperplexitypolicypolysemypositional encodingpost-trainingpower lawPPOpreference alignmentpreference datapretrainingpriorprobabilitypromptprompt engineeringquantizationquery, key, valueRL for reasoningreference policyReLUrepresentational harmresidual streamrewardreward hackingreward modelRLHFRMSNormsamplingscaling lawssegment embedding[SEP] tokensequence classificationsigmoidsoftmaxsparse autoencoderstandard deviationsuperpositiongated FFNsycophancysystem prompttemperaturetest-time computeThe Piletokentop-k samplingtop-ptransfer learningunembeddingvectorvocabularyweight tyingweightsword senseword-sense disambiguationzero-shot