Skip to content

3.0.0-rc3

Pre-release
Pre-release

Choose a tag to compare

@andrewdalpino andrewdalpino released this 18 Sep 02:11
· 121 commits to master since this release
245404d
  • Integers are now considered a categorical data type
  • K Nearest Neighbors and KNN Regressor inference is now parallelized
  • Isolation Forest training and inference is now parallelized
  • Added disk-based streaming neural network snapshotting
  • Added validation interval parameter to MLPs and GBM Learners
  • Cross Entropy loss function now split into Binary and Multiclass
  • Logistic Regression, Softmax, and Adaline now use hold out set
  • Adaboost now uses validation set with early stopping window
  • Renamed TF-IDF dampening parameter to sublinear
  • Exportable Extractors now append by default with option to overwrite
  • RBX Serializer tracks major library version number, not minor
  • Added Class/Cluster Purity clustering metrics
  • V-measure, Completeness, and Homogeneity now use entropy-based formula
  • Fixed KDTree edge pruning + optimize traversal
  • Ball and Vantage Trees now require Subadditive kernels
  • K-d Trees now require Monotonic distance kernels
  • Optimize Dataset sort(), sorting is now unstable
  • Added per-class smoothing to Gaussian Naive Bayes
  • Added per-cluster smoothing to Gaussian Mixture
  • You can now exclude certain categories from one-hot encoding
  • Fixed SVC save/load using class map sidecar
  • Parallel Backends now default to max physical cores not logical
  • Added workers() method to the Backend interface
  • No longer save/load Backend state, transient per environment
  • Added Float Type Converter numeric string and ints to float
  • Boolean Converter now converts truthy and falsy
  • Interval Discretizer now encodes values as integers
  • Polynomial Expander now limited to 10'th degree
  • Updated to PSR-3 Log version 3
  • Update Amp Backend to Amp version 2.0
  • Removed Word Stemmer tokenizer
  • Removed window early stopping from TSNE
  • Removed output layer L2 Penalty parameter from MLP Learners
  • RBX serializer now emits warning on class revision mismatch
  • Class revision hash now compensates for circular references
  • Filesystem Persister now does atomic writes
  • Added cleanup() method to remove neural network residual state
  • Murmur3 new default Token Hashing Vectorizer hash function
  • Dataset fold() now returns excess samples in last fold
  • Increase default Decision Tree max leaf node size from 3 to 5
  • Canonicalized He initializer
  • Xavier 2 now extends He as a deprecated alias
  • Remove Softmax activation function
  • Fix Multiclass layer gradient for non-Cross Entropy losses
  • Optimizers now take a Scheduler rather than a raw learning rate
  • MLP Learners now have gradient accumulation and clipping
  • MLP Learners can now freeze first k layers for fine-tuning
  • PlusPlus and KMC2 Seeders always return unique seeds
  • KMeans and Fuzzy C Means now restrict non-Euclidean kernels
  • NDJSON exporter now preserves zero decimal numbers as floats
  • Add Dataset chunked() factory for online training
  • Added L1 penalty to Dense layers
  • Adaline, Logistic Regression, Softmax Classifier now elastic net
  • Added Iterative interface with progress() method
  • Iterative Learners renamed steps() method to progress()
  • Changed default gradient-based minChange from 1e-4 to 1e-5
  • Prior is now default Strategy of Missing Data Imputer
  • DBSCAN is now a Learner, Probabilistic and Persistable
  • Grid Search now has a fromNamedParams() factory method
  • Grid Search now generates a results() table