495 Commits

Author SHA1 Message Date
Dingyuan Wang
626b415152 fix dict.itervalues mistake 2014-09-07 19:21:13 +08:00
Dingyuan Wang
6a3f228c72 fix python3 stuff 2014-09-07 18:50:10 +08:00
Dingyuan Wang
b16cf0d63f fix indent typo 2014-09-06 23:37:54 +08:00
Dingyuan Wang
6fad5fbb2c update to v0.33 2014-09-06 23:28:47 +08:00
Sun Junyi
fc511de012 Merge pull request #176 from fukuball/master
更新 jieba 可以切換 idf 語料庫及 stop words 語料庫的說明
2014-09-01 14:11:00 +08:00
Sun Junyi
99ea59e88d Update README.md v0.33 2014-08-31 20:04:02 +08:00
fxsjy
6eb43acc10 pip install jieba3k 2014-08-31 20:01:54 +08:00
fxsjy
40adb1c591 version 0.33 2014-08-31 19:26:26 +08:00
Fukuball Lin
d432789cb4 fix typo 2014-08-06 17:56:05 +08:00
Fukuball Lin
cf31a99bf6 將 Readme 中文和半形的英文、數字、符號之間插入空白
將 Readme 中文和半形的英文、數字、符號之間插入空白,增加可讀性
2014-08-06 15:53:57 +08:00
Fukuball Lin
e4d323c78b 更新 jieba 可以切換 idf 語料庫及 stop words 語料庫的說明
更新 jieba 可以切換 idf 語料庫及 stop words 語料庫的說明
2014-08-06 15:00:07 +08:00
Sun Junyi
16d626d347 Merge pull request #174 from fukuball/master
讓 jieba 可以切換 idf 語料庫及 stop words 語料庫
2014-08-06 10:36:10 +08:00
Fukuball Lin
b658ee69cb 讓 jieba 可以自行增加 stop words 語料庫
1. 增加範例 stop words 語料庫
2. 為了讓 jieba 可以切換 stop words 語料庫,新增 set_stop_words 方法,並改寫 extract_tags
3. test 增加 extract_tags_stop_words.py 測試範例
2014-08-06 03:35:16 +08:00
Fukuball Lin
7198d562f1 讓 jieba 可以切換 idf 語料庫
1. 新增繁體中文 idf 語料庫
2. 為了讓 jieba 可以切換 iff 語料庫,新增 get_idf, set_idf_path 方法,並改寫 extract_tags
3. test 增加 extract_tags_idfpath
2014-08-05 22:55:13 +08:00
Sun Junyi
91e5b26f5f Merge pull request #165 from gumblex/jieba3k
fix the u'xxx' string.
2014-06-22 10:23:58 +08:00
Dingyuan Wang
8b07bce568 fix the u'xxx' string. 2014-06-21 23:30:06 +08:00
Sun Junyi
0d99ebce54 Merge pull request #164 from gumblex/jieba3k
Jieba3k v0.32 update
2014-06-15 19:14:28 +08:00
Dingyuan Wang
c04ccd0d12 Update to v0.32 according to the master branch. 2014-06-14 22:31:13 +08:00
Dingyuan Wang
81f77d7a08 Fix the re in enable_parallel. 2014-06-14 15:22:13 +08:00
Sun Junyi
473ac1df75 Merge pull request #162 from ShuraChow/master
fix issue #161
2014-06-11 17:04:23 +08:00
ShuraChow
7583f7760a fix issue #161
posseg每次根据jieba.user_word_tag_tab的长度判断是否有新词载入,如果有,则更新word_tag_tab,然后清空jieba.user_word_tag_tab
2014-06-10 02:04:09 +08:00
Sun Junyi
2726a7c89b Merge pull request #158 from davidlihm/patch-1
Thanks
2014-05-15 10:11:03 +08:00
davidlihm
5b2ec920ed Update __init__.py 2014-05-15 07:55:11 +08:00
Sun Junyi
5574304a9e Merge pull request #152 from jagt/jieba3k
close cache file to avoid warning message.
2014-04-29 11:16:41 +08:00
jagt
7f3513edb7 close cache file to avoid warning message. 2014-04-24 00:35:09 +08:00
Sun Junyi
28621e8b00 Update README.md 2014-04-17 13:47:47 +08:00
Sun Junyi
1f144ebf55 Merge pull request #141 from windch/jieba3k
use logging instead of print in __init__ file of py3k branch
2014-03-20 10:27:52 +08:00
wind
7488b114e7 use logging instead of print in init file 2014-03-20 13:48:33 +13:00
fxsjy
2682e887b8 Merge branch 'master' of https://github.com/fxsjy/jieba 2014-03-02 17:52:52 +08:00
fxsjy
9d4ac26f16 fix the bug of issue#137 2014-03-02 17:52:19 +08:00
Sun Junyi
6942795fae Merge pull request #135 from aszxqw/patch-1
add nodejieba into README.md
2014-02-26 14:13:00 +08:00
Yanyi Wu
ccfa54530e add nodejieba into README.md
add nodejieba into README.md
2014-02-26 14:05:13 +08:00
Sun Junyi
3e430e9769 Update __init__.py v0.32 2014-02-16 20:09:57 +08:00
Sun Junyi
6946b00f14 Merge pull request #134 from Honghe/master
Fix a bug about can not import ChineseAnalyzer
2014-02-16 20:08:42 +08:00
Honghe Wu
7720fbc1d8 fix a bug about can not import ChineseAnalyzer with change tab to 4 wihte spaces under PEP8 2014-02-15 19:32:29 +08:00
fxsjy
cc708de40c version 0.32 released 2014-02-07 15:22:53 +08:00
fxsjy
dafc73425e fix a little problem of dict.txt 2014-02-07 14:35:38 +08:00
fxsjy
7cc7e70843 Merge branch 'master' of https://github.com/fxsjy/jieba 2014-01-28 13:48:35 +08:00
fxsjy
18678d50c6 fix bug issue #132 2014-01-28 13:48:03 +08:00
Sun Junyi
62240c5add Merge pull request #131 from aholic/master
better indent
2014-01-25 18:17:50 -08:00
aholic
e2c796088f better indent 2014-01-24 00:43:48 +08:00
fxsjy
5e6a2c4661 fix a bug of add_word 2013-12-05 13:35:40 +08:00
fxsjy
136676381a fix a bug of add_word 2013-12-05 13:33:24 +08:00
Sun Junyi
e79d54b380 Merge pull request #114 from hermanschaaf/patch-1
Fix typo in error message
2013-10-23 03:41:20 -07:00
Herman Schaaf
95286b8887 Fix typo in error message 2013-10-21 22:21:09 +09:00
fxsjy
14a0ab0466 fix a bug in issue #111 2013-10-11 13:05:59 +08:00
fxsjy
759e1029c8 add an API to control log level: jieba.setLogLevel 2013-09-22 10:26:33 +08:00
Sun Junyi
2ef9dd3a70 Merge pull request #107 from mozillazg/logging
use logging instead of print
2013-09-21 18:54:34 -07:00
Mozillazg
1cf3f0d00b use logging instead of print 2013-09-19 10:31:44 +08:00
Sun Junyi
fd96527f71 Merge pull request #106 from jannson/master
add better support for english for ChineseAnalyzer
2013-09-16 23:58:46 -07:00