index
:
yacy
master
matthew/1.941
upstream/1.941
YaCy search server with some extra patches
summary
refs
log
tree
commit
diff
log msg
author
committer
range
path:
root
/
htroot
/
CrawlStartExpert.html
Age
Commit message (
Expand
)
Author
2025-08-04
more prominent place for collections in crawl start to promote correct
Michael Peter Christen
2023-01-16
added canonical filter
Michael Peter Christen
2023-01-15
front-end integration of tag valency
Michael Peter Christen
2022-09-29
new link to crawlstart api documentation
Michael Peter Christen
2021-08-03
fixed doku link
Michael Peter Christen
2020-12-19
enabling all crawl profiles in all network modes
Michael Peter Christen
2020-01-16
enhanced crawl start url check experience
Michael Peter Christen
2019-05-01
New optional crawl filter on the URL a doc must match to crawl its links
luccioman
2018-10-25
Added new crawler attribute for finer control over Media Type detection
luccioman
2018-10-16
Added a crawl start hint message on availability or not of wkhtmltopdf
luccioman
2018-07-06
Added and updated hint messages about remote crawler status
luccioman
2018-06-19
Added a new crawler document filter type using Solr syntax
luccioman
2018-03-23
Added a crawl filtering possibility on documents Media Type (MIME)
luccioman
2018-03-10
added nav filter
Michael Peter Christen
2018-02-16
Fixed CrawlStartExpert.html HTML validation errors
luccioman
2018-02-16
Issue #156 : new option to clean up (or not) search cache on crawl start
luccioman
2018-02-10
Fixed issue #158 : completed div CSS class ignore in crawl
luccioman
2017-12-19
Updated links to Java Regular Expressions documentation to version 8
luccioman
2017-12-09
added a crawl filter based on <div> tag class names
Michael Peter Christen
2017-06-17
Limit the number of initially previewed links in crawl start pages.
luccioman
2016-11-14
Updated Pattern JavaDoc links to current minimum (1.7) JDK version.
luccioman
2016-11-12
Converted one more set of URLs to pure relative ones.
luccioman
2015-05-08
added must-not-match filter to snapshot generation.
Michael Peter Christen
2015-04-15
enhanced timezone managament for indexed data:
Michael Peter Christen
2015-02-04
remove remote indexing option in crawl start if not in p2p mode
Michael Peter Christen
2015-01-30
added a html field scraper which reads text from html entities of a
Michael Peter Christen
2014-12-09
enhanced the snapshot functionality:
Michael Peter Christen
2014-12-02
get cloned crawl start parameter for snapshots
Michael Peter Christen
2014-12-01
YaCy can now create web page snapshots as pdf documents which can later
Michael Peter Christen
2014-08-27
added hint to the regular expression tester
orbiter
2014-07-18
added an option to set 'obey nofollow' for links with rel="nofollow"
Michael Peter Christen
2014-06-27
fixed external link
Michael Peter Christen
2014-05-17
fix: allow enable of CrawlStartExpert.html #file
reger
2014-04-30
use submitted default userAgent if cloning a crawl
Michael Peter Christen
2014-03-31
made crawl start pages public since they do not reveal individual
orbiter