searxng

mirror of https://github.com/searxng/searxng.git synced 2024-11-05 04:40:11 +01:00

Author	SHA1	Message	Date
Alexandre Flament	436d366448	Merge pull request #2544 from mrwormo/congresslibrary [Engine] Add Library of Congress engine	2021-02-10 10:13:46 +01:00
Alexandre Flament	d2dac11392	[mod] duckduckgo engine: better support of the language preference After the main request, send a second to https://duckduckgo.com/t/sl_h See https://github.com/searx/searx/issues/2259	2021-02-09 14:36:43 +01:00
Markus Heiser	bc1be3f0e9	[enh] add engine MediathekViewWeb (API) Signed-off-by: Markus Heiser <markus.heiser@darmarit.de>	2021-02-09 13:08:01 +01:00
mrwormo	051da88328	Add Library of Congress engine	2021-02-09 12:45:39 +01:00
Alexandre Flament	5e055b069b	[fix) fix apk_mirror engine	2021-02-09 11:02:12 +01:00
Marc Abonce Seguin	64e81794fe	add support for Chinese variants in Wikipedia	2021-02-08 21:56:45 -07:00
Hermógenes Oliveira	514faa9162	[feat] recoll: paged json support	2021-02-07 10:05:35 -03:00
mrwormo	c4c1636b18	Add Creative Commons search engine	2021-02-04 11:31:35 +01:00
Alexandre Flament	ca93a01844	[mod] dynamically set language_support variable The language_support variable is set to True by default, and set to False in only 5 engines. Except the documentation and the /config URL, this variable is not used. This commit remove the variable definition in the engines, and set value according to supported_languages length: False when the length is 0, True otherwise. Close #2485	2021-02-01 17:10:37 +01:00
Markus Heiser	7f505bdc6f	[fix] google: avoid unnecessary SearxEngineXPathException errors Avoid SearxEngineXPathException errors when parsing non valid results:: .//div[@class="yuRUbf"]//a/@href index 0 not found Traceback (most recent call last): File "./searx/engines/google.py", line 274, in response url = eval_xpath_getindex(result, href_xpath, 0) File "./searx/searx/utils.py", line 608, in eval_xpath_getindex raise SearxEngineXPathException(xpath_spec, 'index ' + str(index) + ' not found') searx.exceptions.SearxEngineXPathException: .//div[@class="yuRUbf"]//a/@href index 0 not found Signed-off-by: Markus Heiser <markus.heiser@darmarit.de>	2021-01-28 10:08:50 +01:00
Markus Heiser	b1fefec40d	[fix] normalize the language & region aspects of all google engines BTW: make the engines ready for search.checker: - replace eval_xpath by eval_xpath_getindex and eval_xpath_list - google_images: remove outer try/except block Signed-off-by: Markus Heiser <markus.heiser@darmarit.de>	2021-01-28 10:08:46 +01:00
Markus Heiser	8cdad5d85d	[fix] google-videos: parse values for 'length' & 'author' The 'video.html' template from the 'oscar' design supports replacement for author and length. Google-videos does not have an author, alternatively the publisher info from is used for the author. Hint: these replacements are not supported by the 'simple' design. Signed-off-by: Markus Heiser <markus.heiser@darmarit.de>	2021-01-24 09:51:24 +01:00
Markus Heiser	89b3050b5c	[fix] revise of the google-Video engine This revise is based on the methods developed in the revise of the google engine (see commit `410c2f9`). Signed-off-by: Markus Heiser <markus.heiser@darmarit.de>	2021-01-24 09:39:30 +01:00
Alexandre Flament	8c46b767d0	[fix] google_news: avoid one HTTP redirect except for the English results also add params['soft_max_redirects'] = 1 to avoid false error reporting in /stats/errors	2021-01-24 08:53:35 +01:00
Markus Heiser	5f92dfcdbe	[fix] google-news: query uses locale without country tag Wthout country-region tag google will redirect to correct the contry tag [1]: SEARX_DEBUG=1 searx-checker -v "google news" ... https://news.google.com:443 "GET /search?q=computer&hl=en... HTTP/1.1" 302 0 https://news.google.com:443 "GET /search?q=computer&hl=en-US&.... HTTP/1.1" 200 None ... [1] https://github.com/searx/searx/pull/2483#issuecomment-765600849 Signed-off-by: Markus Heiser <markus.heiser@darmarit.de>	2021-01-23 11:37:14 +01:00
Markus Heiser	baec54c492	[fix] revise of the google-news engine This revise is based on the methods developed in the revise of the google engine (see commit `410c2f9`). Signed-off-by: Markus Heiser <markus.heiser@darmarit.de>	2021-01-22 18:49:45 +01:00
Alexandre Flament	b405646749	Merge pull request #2451 from mrwormo/invidious-engine [Fix] Invidious Engine	2021-01-16 19:25:45 +01:00
Alexandre Flament	a4dcfa025c	[enh] engines: add about variable move meta information from comment to the about variable so the preferences, the documentation can show these information	2021-01-14 20:57:17 +01:00
mrwormo	2dff3887f0	[fix] Invidious engine by enabling requests by randomly picking amongst working instances	2021-01-14 12:12:56 +01:00
Alexandre Flament	3f8ebf70b1	[fix] pylint: use "raise ... from ..."	2020-12-20 09:46:53 +01:00
Alexandre Flament	eb33ae6893	[fix] Python 3.9: use html.unescape instead of HTMLParser.unescape	2020-12-20 09:46:53 +01:00
Alexandre Flament	02fc4147ce	[mod] dictzone, translated, currency_convert: use engine_type online_curency and online_dictionnary	2020-12-17 11:39:36 +01:00
Alexandre Flament	7ec8bc3ea7	[mod] split searx.search into different processors see searx.search.processors.abstract.EngineProcessor First the method searx call the get_params method. If the return value is not None, then the searx call the method search.	2020-12-17 11:39:36 +01:00
lucky13820	fea8958e99	Fix the StartPage result title is showing the url Fix the issue 2395 where StartPage result title is showing the url. https://github.com/searx/searx/issues/2395	2020-12-16 13:54:14 -08:00
Alexandre Flament	292b73a3fc	Merge pull request #2385 from joshu9h/patch-1 [Fix] Startpage	2020-12-14 17:56:48 +01:00
Alexandre Flament	36600118fb	Merge pull request #2372 from dalf/remove-broken-engines [remove] remove searchcode_doc and twitter	2020-12-13 17:11:05 +01:00
joshu9h	8260435c8b	[Fix] Startpage	2020-12-13 15:43:50 +01:00
Alexandre Flament	3c4a9c1188	Merge pull request #2358 from dalf/fix-command [fix] command engine: SearchQuery.query is str not bytes	2020-12-11 14:53:24 +01:00
Alexandre Flament	d703119d3a	[enh] add raise_for_httperror check HTTP response: * detect some comme CAPTCHA challenge (no solving). In this case the engine is suspended for long a time. * otherwise raise HTTPError as before the check is done in poolrequests.py (was before in search.py). update qwant, wikipedia, wikidata to use raise_for_httperror instead of raise_for_status	2020-12-11 14:37:08 +01:00
Alexandre Flament	033f39bff7	Merge pull request #2376 from dalf/fix-mojeek Fix mojeek	2020-12-11 13:14:54 +01:00
Alexandre Flament	6bc6d5e9fd	Merge pull request #2371 from dalf/mod-genius [mod) genious: return valid results even if contents are empty	2020-12-11 13:14:03 +01:00
Alexandre Flament	d41cafd5f3	[fix] xpath, mojeek: fix commit `58d72f2692` before commit `58d72f2`, category was not set in xpath.py, so searx/engines/__init__py was setting the category to ['general'] the commit `58d72f2` set the category to [] which is not replaced by searx/engines/__init__.py consequence: the mojeek engine is hidden in the preferences. this commit revert the xpath.py change. close #2368	2020-12-10 10:52:06 +01:00
Noémi Ványi	3a63dfbdd7	display if an engine does not support https Closes #302	2020-12-09 20:49:54 +01:00
Alexandre Flament	1c9e7cef50	[remove] remove searchcode_doc and twitter * twitter: the API has changed. the engine needs to rewritten. * searchcode_doc: the API about documentation doesn't exist anymore.	2020-12-09 13:14:31 +01:00
Alexandre Flament	fa73f10f11	[mod) genious: return valid results even if contents are empty	2020-12-09 13:01:34 +01:00
Alexandre Flament	a77d8c8227	Merge pull request #2359 from dalf/update-duden [mod] duden engine	2020-12-08 20:33:38 +01:00
Alexandre Flament	bd4869ecd0	Merge pull request #2366 from dalf/remove-seedpeer [remove] seedpeer engine	2020-12-08 20:33:23 +01:00
Alexandre Flament	56c64d6b64	[remove] seedpeer engine the website is offline.	2020-12-07 21:02:29 +01:00
Alexandre Flament	c1a9732268	Merge pull request #2364 from dalf/fix-youtube-noapi [fix] youtube_noapi engine	2020-12-07 20:26:00 +01:00
Alexandre Flament	13d3004703	Merge pull request #2365 from dalf/fix-soundcloud [fix] soundclound: accept result without content	2020-12-07 20:25:17 +01:00
Alexandre Flament	62073c0e1d	Merge pull request #2361 from dalf/fix-1x [fix] 1x engine	2020-12-07 20:24:47 +01:00
Alexandre Flament	923bc02c17	Merge pull request #2363 from dalf/fix-wikipedia-minor [fix] wikipedia: minor fix: return no result instead of crash in some very few cases.	2020-12-07 18:33:37 +01:00
Alexandre Flament	deb1bde20d	[fix] soundclound: accept result without content	2020-12-07 17:45:36 +01:00
Alexandre Flament	34df0f7910	[fix] youtube_noapi engine	2020-12-07 17:44:31 +01:00
Alexandre Flament	58d51e082d	[fix] wikipedia: minor fix: return no result instead of crash in some very few cases. In few cases, the JSON results doesn't contains the key 'type'.	2020-12-07 17:42:05 +01:00
Alexandre Flament	4ec810749b	[fix] 1x engine	2020-12-07 15:46:00 +01:00
Alexandre Flament	1e781863fa	[fix] command engine: SearchQuery.query is str not bytes see `c225db45c8`	2020-12-07 10:43:42 +01:00
Alexandre Flament	9bf594cbcf	[mod] duden engine * add params['soft_max_redirects'] = 1 (when there is spelling suggestion) * avoid try..except * use eval_xpath_* functions	2020-12-07 10:31:11 +01:00
Alexandre Flament	a458451d20	Merge pull request #2356 from dalf/fix-ddd [fix] duckduckgo_definitions: fix relative image URL	2020-12-07 10:16:53 +01:00
Alexandre Flament	925bb561a2	Merge pull request #2352 from dalf/no_http Remove HTTP connections as much as possible	2020-12-06 10:18:49 +01:00
Alexandre Flament	28cc644f0a	[fix] duckduckgo_definitions: fix relative image URL ddg returns relative URL to https://duckduckgo.com/	2020-12-06 10:14:09 +01:00
Alexandre Flament	cdceec1cbb	Merge pull request #2354 from dalf/fix-wikipedia [fix] wikipedia engine: don't raise an error when the query is not found	2020-12-04 20:42:45 +01:00
Alexandre Flament	f0054d67f1	[fix] wikipedia engine: don't raise an error when the query is not found Add a new parameter "raise_for_status", set by default to True. When True, any HTTP status code >= 300 raise an exception ( #2332 ) When False, the engine can manage the HTTP status code by itself.	2020-12-04 20:04:39 +01:00
Alexandre Flament	bef2f2efa8	[fix] wikidata: fix crash when the item has no description at all and at least one URL.	2020-12-04 17:17:20 +01:00
Alexandre Flament	244e812f37	[fix] remove searx/engines/filecrop.py (dead code)	2020-12-04 16:48:15 +01:00
Alexandre Flament	fa909c7c02	[mod] stackoverflow & yandex: detect CAPTCHA response	2020-12-03 13:23:19 +01:00
Alexandre Flament	64cccae99e	[mod] various engines: use eval_xpath* functions and searx.exceptions.* Engine list: ahmia, duckduckgo_images, elasticsearch, google, google_images, google_videos, youtube_api	2020-12-03 10:22:48 +01:00
Alexandre Flament	ad72803ed9	[mod] xpath, 1337x, acgsou, apkmirror, archlinux, arxiv: use eval_xpath_* functions	2020-12-03 10:22:48 +01:00
Alexandre Flament	de887c6347	[mod] bing_news: use eval_xpath_getindex remove unused function searx.utils.list_get	2020-12-03 10:22:48 +01:00
Alexandre Flament	1d0c368746	[enh] record details exception per engine add an new API /stats/errors	2020-12-03 10:22:48 +01:00
Markus Heiser	bef185723a	[refactor] digg - improve results and clean up source code - strip html tags and superfluous quotation marks from content - remove not needed cookie from request - remove superfluous imports Signed-off-by: Markus Heiser <markus.heiser@darmarit.de>	2020-12-02 21:54:27 +01:00
Markus Heiser	6b0a896f01	[mod] digg - pylint searx/engines/digg.py Eliminate redundant file names which are tested by test.pylint and ignored by test.pep8 Signed-off-by: Markus Heiser <markus.heiser@darmarit.de>	2020-12-02 20:59:30 +01:00
Markus Heiser	173b744ef0	[fix] digg - the ISO time stamp of published date has been changed Error pattern:: Engines cannot retrieve results: digg (unexpected crash time data '2020-10-16T14:09:55Z' does not match format '%Y-%m-%d %H:%M:%S') Signed-off-by: Markus Heiser <markus.heiser@darmarit.de>	2020-12-02 20:40:12 +01:00
Alexandre Flament	b00d108673	[mod] pylint: numerous minor code fixes	2020-12-01 15:21:19 +01:00
Alexandre Flament	9ed3ee2beb	[mod] wikidata: WDGeoAttribute class: doesn't change the method signature of get_str	2020-12-01 15:21:17 +01:00
Alexandre Flament	3cfef61123	[fix] /stats: report error percentage instead of error count This bug exists since the PR https://github.com/searx/searx/pull/751	2020-12-01 15:07:09 +01:00
Noémi Ványi	4a36a3044d	Add recoll engine (#2325 ) recoll is a local search engine based on Xapian: http://www.lesbonscomptes.com/recoll/ By itself recoll does not offer web or API access, this can be achieved using recoll-webui: https://framagit.org/medoc92/recollwebui.git This engine uses a custom 'files' result template set `base_url` to the location where recoll-webui can be reached set `dl_prefix` to a location where the file hierarchy as indexed by recoll can be reached set `search_dir` to the part of the indexed file hierarchy to be searched, use an empty string to search the entire search domain	2020-11-30 08:35:15 +01:00
M. Efe Çetin	d1f527c3af	Photon API Link Update Via https://photon.komoot.io/	2020-11-27 10:22:28 +03:00
Alexandre Flament	3786920df9	[enh] Add multiple outgoing proxies credits go to @bauruine see https://github.com/searx/searx/pull/1958	2020-11-20 15:29:21 +01:00
Markus Heiser	c71d214b0c	[refactor] deviantart - improve results and clean up source code Devian's request and response forms has been changed. - fixed title - fixed time_range_dict to 'popular--**' - use image from <noscript> if exists - drop obsolete "http to https, remove domain sharding" - use query URL https://www.deviantart.com/search/deviations?page=5&q=foo - add searx/engines/deviantart.py to pylint check (test.pylint) Error pattern:: There DEBUG:searx:result: invalid title: {'url': 'https://www.deviantart.com/ ... Signed-off-by: Markus Heiser <markus.heiser@darmarit.de>	2020-11-14 17:09:56 +01:00
Alexandre Flament	3038052c79	[mod] remove unused import use from searx.engines.duckduckgo import _fetch_supported_languages, supported_languages_url # NOQA so it is possible to easily remove all unused import using autoflake: autoflake --in-place --recursive --remove-all-unused-imports searx tests	2020-11-14 14:11:02 +01:00
Alexandre Flament	c3d9b17c2a	Merge pull request #2292 from kvch/elasticsearch-engine New engine: Elasticsearch	2020-11-14 13:25:08 +01:00
Alexandre Flament	102c08838b	Merge pull request #2289 from dalf/pylint [mod] pylint: add extension-pkg-whitelist=lxml.etree	2020-11-14 13:24:31 +01:00
Noémi Ványi	43e697681e	New engine: Elasticsearch	2020-11-10 19:53:38 +01:00
Alexandre Flament	58d72f2692	[mod] pylint: minor code change to allow pylint globally This commit is only a step, it doesn't fix all the issues reported by pylint	2020-11-03 11:35:53 +01:00
Alexandre Flament	eed43783f9	[fix] comamnd engine: fix import	2020-11-03 10:55:08 +01:00
Alexandre Flament	a08df82574	[fix] scanr_structure engine: fix import	2020-11-03 10:54:02 +01:00
Alexandre Flament	95bd6033fa	[mod] wikidata engine: use one SPARQL request instead of 2 HTTP requests.	2020-10-28 08:09:25 +01:00
Alexandre Flament	ca593728af	[mod] duckduckgo_definitions: display only user friendly attributes / URL various bug fixes	2020-10-28 08:09:25 +01:00
a01200356	c3daa08537	[enh] Add onions category with Ahmia, Not Evil and Torch Xpath engine and results template changed to account for the fact that archive.org doesn't cache .onions, though some onion engines migth have their own cache. Disabled by default. Can be enabled by setting the SOCKS proxies to wherever Tor is listening and setting using_tor_proxy as True. Requires Tor and updating packages. To avoid manually adding the timeout on each engine, you can set extra_proxy_timeout to account for Tor's (or whatever proxy used) extra time.	2020-10-25 17:59:05 -07:00
Nicholas Kegler	8e15d3e4c1	Open Semantic Search Engine	2020-10-25 17:50:00 +01:00
Noémi Ványi	e158eeee4b	Propagate error messages from YouTube API	2020-10-09 17:34:26 +02:00
Adam Tauber	835d16cbb1	Merge pull request #2255 from kvch/yacy-improvements Add yacy improvements: HTTP digest auth, category checking	2020-10-09 16:34:42 +02:00
Alexandre Flament	cfd21bc475	[fix] fix duckduckgo engine - remove paging support: a "vqd" parameter is required between each request. This parameter is uniq for each request - update the URL (no redirect), use the POST method - language support: works if there is no more than request per minute, otherwise it is ignored !	2020-10-09 16:00:42 +02:00
Noémi Ványi	72c7fd25fe	Add yacy improvements: HTTP digest auth, category checking	2020-10-09 15:06:05 +02:00
Noémi Ványi	f0278d41fc	add ebay enginte to shopping category	2020-10-08 13:20:55 +02:00
Alexandre Flament	a9dc54bebc	[mod] Add searx.data module Instead of loading the data/*.json in different location, load these files in the new searx.data module.	2020-10-07 10:29:34 +02:00
Alexandre Flament	8659212f5a	[fix] drop Python 2: use collections.abc.Iterable instead of collections.Iterable	2020-10-06 09:43:24 +02:00
Alexandre Flament	b728cb610b	Merge pull request #2241 from dalf/move-extract-text-and-url Move the extract_text and extract_url functions to searx.utils	2020-10-04 09:06:20 +02:00
Finn	53c8d945b4	[enh] Add SepiaSearch engine (#2227 ) supported_languages values: see https://framagit.org/framasoft/peertube/search-index/-/blob/master/client/src/views/Search.vue#L618-641	2020-10-03 13:00:10 +02:00
Alexandre Flament	2006eb4680	[mod] move extract_text, extract_url to searx.utils	2020-10-02 18:13:56 +02:00
Markus Heiser	8162d7aff4	[fix] google engine - div classes has been renamed in HTML reult Since 1. October 2020 google has changed the 'class' attribute of the HTML result page. Fix the xpath expressions and ignore <div class="g" ../> sections which do not match to title's xpath expression. Signed-off-by: Markus Heiser <markus.heiser@darmarit.de>	2020-10-01 09:44:29 +02:00
Alexandre Flament	f204e4903d	[fix] migration from github.com/asciimoo/searx to github.com/searx/searx : fix URLs	2020-09-28 16:44:14 +02:00
Marc Abonce Seguin	ecf5899153	fetch google's search langs rather than ui langs	2020-09-22 11:37:44 +02:00
Marc Abonce Seguin	41800835f9	fetch supported languages for startpage engine	2020-09-22 11:37:44 +02:00
Marc Abonce Seguin	ea9d979cc3	add language names in qwant's fetch languages function	2020-09-22 11:37:44 +02:00
Dalf	c225db45c8	Drop Python 2 (4/n): SearchQuery.query is a str instead of bytes	2020-09-10 10:49:42 +02:00
Dalf	1022228d95	Drop Python 2 (1/n): remove unicode string and url_utils	2020-09-10 10:39:04 +02:00
Marc Abonce Seguin	ab20ca182c	use Wikipedia's REST v1 API	2020-09-10 09:54:30 +02:00
Noémi Ványi	f0ca1c3483	[enh] Add command line engines: git grep, find, etc. (#2128 ) A new "base" engine called command is introduced. It is the foundation for all command line engines for now. You can use this engine to create your own command line engine. Add some engines (commented out to make sure no one enables anything accidentally): * git grep: This engine lets you grep in the searx repo. * locate: If locate is installed and initialized, you can search on the FS. * find: You can find files with a specific name from where you started searx. * pattern search in files: This engine utilizes the command fgrep. * regex search in files: This engine runs `grep` to find a file based on its contents.	2020-09-08 09:51:53 +02:00

1 2 3 4 5 ...

1094 Commits