Bryan Woods
c3f973b1de
FIX supports 64 bit group ID and indexes ( #12736 )
2019-01-09 09:52:10 +11:00
Roman Yurchak
701144559f
MAINT Remove unused utils.fixes ( #12928 )
...
This continues the work done in https://github.com/scikit-learn/scikit-learn/pull/12639 on dropping the python 2 support by,
- ~~removing unnecessary `from __future__` imports~~
- removing unused `sklearn.utils.fixes` assuming we can agree in https://github.com/scikit-learn/scikit-learn/issues/12927 that `sklearn.utils.fixes` are private as was stated e.g. in https://github.com/scikit-learn/scikit-learn/issues/6616#issuecomment-245109979
2019-01-08 12:43:49 +11:00
Roman Yurchak
acb8106472
MNT Use list and dict comprehension ( #12668 )
2019-01-08 09:32:34 +11:00
Andreas Mueller
952ef6637a
MRG Drop legacy python / remove six dependencies ( #12639 )
2019-01-03 15:50:05 +02:00
Emmanuel Arias
19a7c08c82
DOC Document that fetch_20newsgroups also returns target_names ( #12783 )
2018-12-23 09:55:08 +08:00
Adrin Jalali
2bd87f6ed8
Remove python < 3.5 from CI ( #12746 )
2018-12-14 10:53:12 +01:00
Bartosz Michałowski
fa98a72dcc
MNT Replaced all occurrences of assert_true and assert_false with assert ( #12588 )
2018-11-28 09:16:26 +08:00
Hanmin Qin
12384a1f39
MNT silence a LGTM alert ( #12633 )
2018-11-22 09:03:34 +08:00
Thomas Moreau
d25da1be20
FIX make joblib utils private, and remove mentions of externals.joblib ( #12345 )
2018-11-20 10:53:52 +11:00
Thomas Fan
14c816e2b8
ENH/FIX openml, Adds retrying if reading from cache fails ( #12526 )
2018-11-14 17:40:35 +11:00
Hanmin Qin
43e3a02085
MNT Remove unused assert_true imports ( #12560 )
2018-11-11 11:08:37 +08:00
Yaroslav Halchenko
362cb3bcab
TST autoreplace assert_true(...==...) with plain assert ( #12547 )
2018-11-11 09:05:34 +08:00
janvanrijn
3282d43ccc
[MRG] Additional Warnings in case OpenML auto-detected a problem with dataset ( #12541 )
...
* added additional warning output
* added features gzip
* added gzipped datasets
* fix file naming
* changed expected warning msg
2018-11-07 10:30:33 -05:00
Olivier Grisel
e67e30cec2
Fix numpy vstack on generator expressions ( #12467 )
...
* Workaround vstack issue with genxp
* Use list comprehensions instead of genexps with np.vstack
* Add changelog entry.
2018-10-29 11:40:05 -04:00
jeremiedbb
013d295a13
FIX olivetti_faces DESCR to point to the good location ( #12441 )
2018-10-23 21:43:58 +08:00
Matthew Roeschke
5af272ac61
MNT Remove unused variables ( #12230 )
2018-10-22 12:48:40 +11:00
Andreas Mueller
0f94f2962b
MNT simple deprecations and removals for 0.21 ( #12238 )
...
Part of #11992 .
These were all the things that seemed pretty straight-forward. It's actually a bit bulky but should still be easy to review, hopefully.
2018-10-11 14:56:37 -04:00
janvanrijn
03c3af5bde
[MRG] Fix fetch_openml when ignore attributes are numeric ( #12330 )
...
* modularized data column functionality
* small bugfix
* removes redundant line breaks
* added some documentation on the added fn
* added additional comment on advice of Nicholas Hug
* added test case
* merged master into branch, and added small comments by Joel
* added doc item
2018-10-09 11:25:46 -04:00
TakingItCasual
f4e7d2b19a
Converting http to https (3)... ( #12302 )
2018-10-05 18:50:31 +02:00
janvanrijn
afa0694d12
FIX cache of OpenML fetcher ( #12246 )
2018-10-05 16:36:06 +10:00
TakingItCasual
1e052e9da9
Converting http to https (2)... ( #12292 )
2018-10-04 23:06:14 +02:00
Roman Yurchak
e0e7387606
Remove unused private functions ( #12253 )
2018-10-03 21:08:17 -04:00
Roman Feldbauer
60cf1d62d2
Fix numpy.int overflow in make_classification ( #10811 )
2018-10-02 14:15:56 +02:00
Hanmin Qin
06b4307fbc
DOC Include fetch_openml doc in user guide ( #12065 )
2018-09-13 18:17:31 +08:00
Adrin Jalali
36536c6f46
MAINT Fix invalid escape sequence ( #12064 )
2018-09-13 17:08:23 +08:00
Umar Farouk Umar
a056a57325
DOC fix for linnerud dataset ( #12024 )
...
The descriptions were the wrong way around
2018-09-06 21:10:17 +10:00
Joel Nothman
1fafc5c56d
TST use urlopen monkeypatch for test_decode_* ( #12020 )
...
Avoid requiring internet for test suite. Examples will still run with internet (as long as cache is occasionally cleared).
2018-09-06 09:29:37 +02:00
Gabriele Calvo
e726f7a3e6
DOC fix minor spacing issue in the iris dataset description ( #12019 )
2018-09-05 22:50:20 +02:00
Thomas Fan
83e73759f0
ENH Uses gzip when caching in fetch_openml ( #11830 )
2018-09-03 09:03:44 +10:00
Adrin Jalali
dd4b528612
[MRG+1] fetch_openml: test data file names resemble the urls ( #11846 )
2018-08-21 16:13:06 +02:00
vufg
52a36b4d1a
ENH fetch_openml should support return_X_y ( #11840 )
2018-08-19 22:11:16 +08:00
Joel Nothman
fc56da504d
Deprecate fetch_mldata ( #11466 )
...
* API Deprecate fetch_mldata and update examples
* Use pytest's filterwarnings
* Rm unused import
* Remove broken doctest
* Refer user to openml URL
* DOC whatsnew tweak
2018-08-18 19:27:48 +03:00
Adrin Jalali
d2d9abb48d
ENH fetch_openml: more strongly encourage users to specify version ( #11827 )
2018-08-16 16:07:44 +03:00
janvanrijn
ab82f5739f
[MRG] Openml data loader ( #11419 )
2018-08-15 17:21:20 +10:00
Hanmin Qin
631711a453
DOC incorrect n_samples in fetch_species_distributions
2018-07-27 10:50:21 +08:00
Hanmin Qin
f1c967883b
DOC fetch_20newsgroups_vectorized is based on CountVectorizer ( #11685 )
2018-07-26 12:06:11 +08:00
jeremiedbb
9d649c5e5b
DOC Clean up datasets loaders as part of the reorganization of the dataset section ( #11319 )
...
Standardize the datasets informations, as part of a more general reorganization of the dataset section in user guide, see #11083 .
Fixes #10555
2018-07-25 23:52:02 +10:00
Nicolas Hug
72b2ed9ee1
DOC Added docstring checks for dataset module ( #11407 )
2018-07-24 10:57:13 +10:00
ZJ Poh
adddf00433
[MRG] np.ones -> np.full ( #11628 )
2018-07-23 09:49:01 +02:00
Gael Varoquaux
47d3b3c30f
[MRG] TST: avoid DeprecationWarning on latest joblib ( #11625 )
2018-07-20 16:13:50 +02:00
Ronan Lamy
5592a2eda9
[MRG] PyPy support for all but a couple of estimators ( #11010 )
2018-07-20 14:39:53 +10:00
Joel Nothman
14e7c328df
Restructure access to vendored/site Joblib ( #11471 )
...
In order to fix #11408 , this swaps `joblib` and `_joblib`. It however, allows users to access joblib's `Memory` or `Parallel` functionality without accessing `sklearn.externals._joblib` by importing `Memory`, `Parallel`, etc. into `sklearn.utils`.
2018-07-17 18:02:11 +02:00
Andreas Mueller
b25e222328
MNT misc sphinx website build fixes and dead links ( #11532 )
2018-07-16 09:13:20 +08:00
Roman Yurchak
c09352c241
[MRG] Fix DeprecationWarning due to collections.abc in Python 3.7 ( #11431 )
...
Closes https://github.com/scikit-learn/scikit-learn/issues/11121
This PR removes the deprecation warning about ABC being moved from `collections` to `collections.abc` when importing scikit-learn in Python 3.7.
In the end, I put `collections.abc.{Sequence, Iterable, Mapping, Sized}` in the namespace of `sklearn.utils.fixes`. This was the simplest way I could find, and while it has the drawback of obfuscating the real module name, other approached appeared more problematic and a similar approach is currently used e.g. for `utils.fixes.signature` which is an alias for `inspect.signature`.
We can't just patch six with https://github.com/benjaminp/six/pull/241 , because sklearn uses six from 5 years ago, which would need updating and I'm not sure if it could have side effects (e.g. for pickling backward compatibility etc).
**Edit**: This adds a test checking that generally no warnings are raised when importing scikit-learn top-level modules.
2018-07-14 15:14:02 -05:00
Joel Nothman
14764061f8
[MRG] DOC fix some sphinx warnings ( #11241 )
2018-06-21 20:43:21 +10:00
jeremiedbb
1ff8364387
DOC reorganize datasets documentation page ( #11180 )
2018-06-19 10:40:33 +10:00
Roman Yurchak
d582e97944
TST Pytest parametrization part2 - cluster, datasets and decomposition modules ( #11142 )
2018-06-01 10:26:02 +08:00
Hanmin Qin
399f1b2761
FIX Correct iris dataset ( #11082 )
2018-05-22 12:56:57 +08:00
Taehoon Lee
328b04f43e
MNT Fix typos ( #11057 )
2018-05-03 16:30:05 +08:00
Nicholas Nadeau, P.Eng., AVS
3e26fc63be
MAINT Fixing Typos ( #11017 )
2018-04-24 09:32:25 +10:00