Ranked from CortexLab’s local platform index and published content.
Q-Learning
An off-policy temporal-difference method that learns action values from reward transitions.
Contextual Bandit
Chooses among actions using context while learning from immediate reward.
Multilayer Perceptron
A stack of learned affine transformations and nonlinearities that approximates complex functions.
BM25
A strong lexical ranking algorithm balancing term frequency, rarity and document length.
Thompson Sampling
Samples action quality from posterior beliefs to balance exploration and exploitation.
Graph Neural Network
Learns node/edge representations by passing messages across graph neighborhoods.
Markov Decision Process
A formal model of states, actions, transition probabilities, rewards and discounting.
Logistic Regression
A probabilistic classification baseline that is transparent, fast and surprisingly competitive.