diff options
author | Sean Owen <sowen@cloudera.com> | 2014-07-13 19:27:43 -0700 |
---|---|---|
committer | Xiangrui Meng <meng@databricks.com> | 2014-07-13 19:27:43 -0700 |
commit | 635888cbed0e3f4127252fb84db449f0cc9ed659 (patch) | |
tree | 43433e3393c889f25a8ef4898099664a1a5ce0a7 /docs/mllib-optimization.md | |
parent | 4c8be64e768fe71643b37f1e82f619c8aeac6eff (diff) | |
download | spark-635888cbed0e3f4127252fb84db449f0cc9ed659.tar.gz spark-635888cbed0e3f4127252fb84db449f0cc9ed659.tar.bz2 spark-635888cbed0e3f4127252fb84db449f0cc9ed659.zip |
SPARK-2363. Clean MLlib's sample data files
(Just made a PR for this, mengxr was the reporter of:)
MLlib has sample data under serveral folders:
1) data/mllib
2) data/
3) mllib/data/*
Per previous discussion with Matei Zaharia, we want to put them under `data/mllib` and clean outdated files.
Author: Sean Owen <sowen@cloudera.com>
Closes #1394 from srowen/SPARK-2363 and squashes the following commits:
54313dd [Sean Owen] Move ML example data from /mllib/data/ and /data/ into /data/mllib/
Diffstat (limited to 'docs/mllib-optimization.md')
-rw-r--r-- | docs/mllib-optimization.md | 2 |
1 files changed, 1 insertions, 1 deletions
diff --git a/docs/mllib-optimization.md b/docs/mllib-optimization.md index ae9ede58e8..651958c781 100644 --- a/docs/mllib-optimization.md +++ b/docs/mllib-optimization.md @@ -214,7 +214,7 @@ import org.apache.spark.mllib.linalg.Vectors import org.apache.spark.mllib.util.MLUtils import org.apache.spark.mllib.classification.LogisticRegressionModel -val data = MLUtils.loadLibSVMFile(sc, "mllib/data/sample_libsvm_data.txt") +val data = MLUtils.loadLibSVMFile(sc, "data/mllib/sample_libsvm_data.txt") val numFeatures = data.take(1)(0).features.size // Split data into training (60%) and test (40%). |