Permalink
Please sign in to comment.
Browse files
[SPARK-16421][EXAMPLES][ML] Improve ML Example Outputs
## What changes were proposed in this pull request? Improve example outputs to better reflect the functionality that is being presented. This mostly consisted of modifying what was printed at the end of the example, such as calling show() with truncate=False, but sometimes required minor tweaks in the example data to get relevant output. Explicitly set parameters when they are used as part of the example. Fixed Java examples that failed to run because of using old-style MLlib Vectors or problem with schema. Synced examples between different APIs. ## How was this patch tested? Ran each example for Scala, Python, and Java and made sure output was legible on a terminal of width 100. Author: Bryan Cutler <[email protected]> Closes #14308 from BryanCutler/ml-examples-improve-output-SPARK-16260.
- Loading branch information...
Showing
with
427 additions
and 2,757 deletions.
- +0 −1,000 data/mllib/lr-data/random.data
- +0 −1,000 data/mllib/lr_data.txt
- +0 −569 data/mllib/sample_tree_data.csv
- +5 −0 examples/src/main/java/org/apache/spark/examples/JavaPageRank.java
- +3 −2 examples/src/main/java/org/apache/spark/examples/ml/JavaAFTSurvivalRegressionExample.java
- +6 −5 examples/src/main/java/org/apache/spark/examples/ml/JavaBinarizerExample.java
- +6 −1 examples/src/main/java/org/apache/spark/examples/ml/JavaBucketizerExample.java
- +4 −0 examples/src/main/java/org/apache/spark/examples/ml/JavaChiSqSelectorExample.java
- +1 −1 examples/src/main/java/org/apache/spark/examples/ml/JavaCountVectorizerExample.java
- +5 −1 examples/src/main/java/org/apache/spark/examples/ml/JavaDCTExample.java
- +2 −2 examples/src/main/java/org/apache/spark/examples/ml/JavaGaussianMixtureExample.java
- +14 −1 examples/src/main/java/org/apache/spark/examples/ml/JavaIndexToStringExample.java
- +2 −2 examples/src/main/java/org/apache/spark/examples/ml/JavaIsotonicRegressionExample.java
- +23 −5 examples/src/main/java/org/apache/spark/examples/ml/JavaMaxAbsScalerExample.java
- +25 −5 examples/src/main/java/org/apache/spark/examples/ml/JavaMinMaxScalerExample.java
- +7 −1 examples/src/main/java/org/apache/spark/examples/ml/JavaMultilayerPerceptronClassifierExample.java
- +7 −11 examples/src/main/java/org/apache/spark/examples/ml/JavaNGramExample.java
- +10 −3 examples/src/main/java/org/apache/spark/examples/ml/JavaNaiveBayesExample.java
- +21 −2 examples/src/main/java/org/apache/spark/examples/ml/JavaNormalizerExample.java
- +3 −1 examples/src/main/java/org/apache/spark/examples/ml/JavaOneHotEncoderExample.java
- +1 −1 examples/src/main/java/org/apache/spark/examples/ml/JavaOneVsRestExample.java
- +1 −1 examples/src/main/java/org/apache/spark/examples/ml/JavaPCAExample.java
- +5 −9 examples/src/main/java/org/apache/spark/examples/ml/JavaPolynomialExpansionExample.java
- +1 −1 examples/src/main/java/org/apache/spark/examples/ml/JavaStopWordsRemoverExample.java
- +3 −0 examples/src/main/java/org/apache/spark/examples/ml/JavaStringIndexerExample.java
- +5 −7 examples/src/main/java/org/apache/spark/examples/ml/JavaTfIdfExample.java
- +22 −11 examples/src/main/java/org/apache/spark/examples/ml/JavaTokenizerExample.java
- +4 −2 examples/src/main/java/org/apache/spark/examples/ml/JavaVectorAssemblerExample.java
- +2 −2 examples/src/main/java/org/apache/spark/examples/ml/JavaVectorSlicerExample.java
- +7 −2 examples/src/main/java/org/apache/spark/examples/ml/JavaWord2VecExample.java
- +6 −4 examples/src/main/python/ml/binarizer_example.py
- +3 −1 examples/src/main/python/ml/bucketizer_example.py
- +2 −0 examples/src/main/python/ml/chisq_selector_example.py
- +3 −1 examples/src/main/python/ml/count_vectorizer_example.py
- +1 −2 examples/src/main/python/ml/dct_example.py
- +3 −3 examples/src/main/python/ml/gaussian_mixture_example.py
- +11 −3 examples/src/main/python/ml/index_to_string_example.py
- +2 −2 examples/src/main/python/ml/isotonic_regression_example.py
- +10 −2 examples/src/main/python/ml/linear_regression_with_elastic_net.py
- +8 −2 examples/src/main/python/ml/max_abs_scaler_example.py
- +8 −2 examples/src/main/python/ml/min_max_scaler_example.py
- +1 −1 examples/src/main/python/ml/multilayer_perceptron_classification.py
- +4 −5 examples/src/main/python/ml/n_gram_example.py
- +8 −4 examples/src/main/python/ml/naive_bayes_example.py
- +8 −1 examples/src/main/python/ml/normalizer_example.py
- +2 −2 examples/src/main/python/ml/onehot_encoder_example.py
- +3 −2 examples/src/main/python/ml/pipeline_example.py
- +5 −6 examples/src/main/python/ml/polynomial_expansion_example.py
- +1 −1 examples/src/main/python/ml/stopwords_remover_example.py
- +4 −5 examples/src/main/python/ml/tf_idf_example.py
- +9 −5 examples/src/main/python/ml/tokenizer_example.py
- +4 −3 examples/src/main/python/ml/train_validation_split.py
- +2 −1 examples/src/main/python/ml/vector_assembler_example.py
- +4 −0 examples/src/main/python/ml/vector_indexer_example.py
- +3 −2 examples/src/main/python/ml/word2vec_example.py
- +5 −2 examples/src/main/python/pagerank.py
- +5 −0 examples/src/main/scala/org/apache/spark/examples/SparkPageRank.scala
- +3 −2 examples/src/main/scala/org/apache/spark/examples/ml/AFTSurvivalRegressionExample.scala
- +5 −3 examples/src/main/scala/org/apache/spark/examples/ml/BinarizerExample.scala
- +4 −1 examples/src/main/scala/org/apache/spark/examples/ml/BucketizerExample.scala
- +3 −0 examples/src/main/scala/org/apache/spark/examples/ml/ChiSqSelectorExample.scala
- +1 −1 examples/src/main/scala/org/apache/spark/examples/ml/CountVectorizerExample.scala
- +1 −1 examples/src/main/scala/org/apache/spark/examples/ml/DCTExample.scala
- +2 −2 examples/src/main/scala/org/apache/spark/examples/ml/GaussianMixtureExample.scala
- +13 −1 examples/src/main/scala/org/apache/spark/examples/ml/IndexToStringExample.scala
- +2 −2 examples/src/main/scala/org/apache/spark/examples/ml/IsotonicRegressionExample.scala
- +1 −1 examples/src/main/scala/org/apache/spark/examples/ml/LinearRegressionWithElasticNetExample.scala
- +2 −1 examples/src/main/scala/org/apache/spark/examples/ml/LogisticRegressionSummaryExample.scala
- +8 −2 examples/src/main/scala/org/apache/spark/examples/ml/MaxAbsScalerExample.scala
- +8 −2 examples/src/main/scala/org/apache/spark/examples/ml/MinMaxScalerExample.scala
- +1 −1 examples/src/main/scala/org/apache/spark/examples/ml/MultilayerPerceptronClassifierExample.scala
- +4 −3 examples/src/main/scala/org/apache/spark/examples/ml/NGramExample.scala
- +1 −1 examples/src/main/scala/org/apache/spark/examples/ml/NaiveBayesExample.scala
- +8 −1 examples/src/main/scala/org/apache/spark/examples/ml/NormalizerExample.scala
- +2 −1 examples/src/main/scala/org/apache/spark/examples/ml/OneHotEncoderExample.scala
- +1 −1 examples/src/main/scala/org/apache/spark/examples/ml/OneVsRestExample.scala
- +4 −3 examples/src/main/scala/org/apache/spark/examples/ml/PCAExample.scala
- +7 −5 examples/src/main/scala/org/apache/spark/examples/ml/PolynomialExpansionExample.scala
- +1 −1 examples/src/main/scala/org/apache/spark/examples/ml/StopWordsRemoverExample.scala
- +4 −4 examples/src/main/scala/org/apache/spark/examples/ml/TfIdfExample.scala
- +8 −3 examples/src/main/scala/org/apache/spark/examples/ml/TokenizerExample.scala
- +2 −0 examples/src/main/scala/org/apache/spark/examples/ml/UnaryTransformerExample.scala
- +2 −1 examples/src/main/scala/org/apache/spark/examples/ml/VectorAssemblerExample.scala
- +5 −2 examples/src/main/scala/org/apache/spark/examples/ml/VectorSlicerExample.scala
- +4 −1 examples/src/main/scala/org/apache/spark/examples/ml/Word2VecExample.scala
Oops, something went wrong.
0 comments on commit
180fd3e