Commit Graph

499 Commits (5ff979405d3255568a53b24fc242f68dfb5979da)

Author SHA1 Message Date
orbiter b21b9df2d0 added section headlines generation to html parser
19 years ago
rramthun c4487deba9 Minor changes collected over some time.
19 years ago
allo 6822dce57b Using Orbiters function for auth
19 years ago
orbiter 2028403670 - consolidated different orderings to kelondroNaturalOrder
19 years ago
orbiter 9544c47684 added some UTF-8 handling.
19 years ago
borg-0300 9d8dca750e BUGFIX for my last commit
19 years ago
borg-0300 5449193167 bugfix for http://www.yacy-forum.de/viewtopic.php?t=1706 (i hope)
19 years ago
borg-0300 2a23f5d419 F..., Sorry, no time, later
19 years ago
borg-0300 3a2d13786e bugfix for http://www.yacy-forum.de/viewtopic.php?t=1706
19 years ago
borg-0300 dc0999ec9c adapted to new HTCache structure
19 years ago
orbiter 9086261476 refactoring of base64 encoding:
19 years ago
borg-0300 b24fcc8ca4 oom
19 years ago
borg-0300 7da232b5b9 HTCache Reset if necessary
19 years ago
borg-0300 4f18f24d81 small change
19 years ago
borg-0300 c652527620 YaCy removes now the old HTCACHE data
19 years ago
borg-0300 69f65210e2 ".yacy" has its own directory;
19 years ago
allo 351fffc129 DATA/WORK for user-created content
19 years ago
allo a81cc9d969 no DATA/DATA to avoid confusion.
19 years ago
borg-0300 b95c5d5781 BUGFIX for URLs how "/../" ...;
19 years ago
allo 9cce3c5709 dates Table for bookmarksdb(needed for del.icio.us api)
19 years ago
hermens 11fe95832e avoid division by zero when index transfer is extremely fast
19 years ago
allo 4ac0fd328a First Version of the Bookmarksmanager
19 years ago
theli d7b6dcbe2e *) Bugfix for MalformedURL problem if Location header is empty.
19 years ago
hermens 5b3e01bd3c avoid division by zero when importing very small indexes (<100 entries)
19 years ago
borg-0300 b7f9adc2c9 new filters added
19 years ago
theli 79667a172e *) Bugfix for additional parser problem
19 years ago
theli 8c594841a8 *) Bugfix for incorrectly indexing of URLs that were requested with Cookies in the
19 years ago
orbiter b5d02d649a fixed bug caused strange search result behaviour
19 years ago
orbiter 4500506735 fixed some bugs concerning url entry retrieval and intexControl interface
19 years ago
orbiter 83a34b838d * added Object allocation monitor on performanceMemory page
19 years ago
orbiter 4ff3d219e8 increased delay for cacheScan start and slowed down scan process
19 years ago
orbiter 3031903d50 re-design of RAM cache flush into assortment cluster
19 years ago
orbiter 0c762daf4b better startup failure handling
19 years ago
orbiter f27f9ecf15 * activated write buffer for databases.
19 years ago
orbiter c59d1b2f5e - Tests with write buffer (new class kelondroBufferedIOChunks, not yet active)
19 years ago
orbiter bb79fb5d91 - changed handling of error cases retrieving urls from database
19 years ago
theli e7d16ef831 *) Corrections in jMimeMagic MagicRule-file to detect some special rss feeds
19 years ago
theli 386d9e45d8 *) Bugfix for code cleanup
19 years ago
theli 5a1d45715d *) Bugfix for parser configuration bug
19 years ago
rramthun a1061495d4 Fixed some spelling mistakes and added some text which (should) make it easier to understand the options.
19 years ago
orbiter 0cdc58aaea fixed indexing of local domains.
19 years ago
theli e1c2d8ec5f *) Speedup "removed from queue"
19 years ago
hydrox 96930f0d2b *)added function to removed malformed URLs from urlHash.db
19 years ago
theli 8862b6ba4b *) Corrections for code cleanup 1175
19 years ago
orbiter 13fdebc50d added authentication for link deletion in search result
19 years ago
orbiter 37f88b4017 code cleanup
19 years ago
orbiter ec2b39c1ce code cleanup
19 years ago
orbiter 8f1f2daa5e implemented interactive link deletion of search results.
19 years ago
theli 6d0f7e6988 *) Adding missing file
19 years ago
theli 44fa94ac52 *) Modifications for dbImport functionality
19 years ago
orbiter dc778659fb fixed problem with time-out during result joint which caused OR behavior instead of AND beahvior
19 years ago
orbiter 3d8a5ae652 code cleanup
19 years ago
theli 64478b1f02 *) Adding possibility to delete crawler queue entries using regular expressions
19 years ago
orbiter a04930f025 code cleanup
19 years ago
low012 90b0eb144e just a typo...
19 years ago
theli 129b15f3e1 *) Correcting logging output of db importer thread
19 years ago
orbiter 420d56ce79 extended db-testing
19 years ago
orbiter ecf765ec33 temporary fix to make jrpm extension compilable with my netbeans environment
19 years ago
theli 8ed0aaae8d *) Adding content Parser for RPM Files
19 years ago
theli 818d37ce44 *) Removing getSimpleName
19 years ago
theli b35c5a48bf *) First version of urlRedirector.pl script
19 years ago
theli bdf30117c1 *) Redesign of parser configuration
19 years ago
theli d4ac3e25b1 *) Bugfix for file system link bug during detection of invalid URLs
19 years ago
orbiter adf75bc9fa better logging for invalid file path detection
19 years ago
orbiter 40621a5663 anhancements in ranking preparation and fixed problem with parser/mime recognition
19 years ago
theli c650b112ea *) Bugfix for relative URL Bug in Crawler
19 years ago
theli 4e73035aef *) Bugfix for "too many open files" during index distribution
19 years ago
orbiter f57e2d67f5 shortened network overview (less columns fit easier on page)
19 years ago
orbiter 85282b1d98 enhanced YBR recognition and search result heuristics
19 years ago
orbiter b9cc9029e3 added ybr selection for remote search
19 years ago
orbiter 0e25020f51 added first generation and usage of YBR index-files. Enhanced overall ranking of search results.
19 years ago
theli 90d6c6223b *) Adding color codes to network graphic legend
19 years ago
orbiter bfe51c7228 added generation of domain-list
19 years ago
orbiter 0ec54d9c5f enhanced CR-file handling and added first RCI-evaluation tests
19 years ago
theli c2fe3a1670 *) Updating jMimeMagic Ruleset
19 years ago
orbiter 88e3234393 fine-tuning of rci-generation
19 years ago
orbiter a12759c1bf first try to implement a rci-computation from cr-files
19 years ago
orbiter 4a8e8f269e refactoring of cr-processing; new kelondro class to handle the attribute file format
19 years ago
orbiter 24dc0e0760 implemented cr-file processing and further transmission steps
19 years ago
orbiter 9d9a87f445 limited htcache storage length
19 years ago
theli d0dfccdb77 *) Making CrawlStacker pool configurable via GUI and config file
19 years ago
theli 3631cb1f6d *) deleting empty entities during index selection
19 years ago
theli ca26aab9b1 *) More debugging output for migrateWords
19 years ago
theli 9b35ae9027 *) Correcting wrong % values on IndexTransfer_p page
19 years ago
theli e6bf9d90a5 *) Fixing Problems with MalformedURLs during Word Selection
19 years ago
theli 86a9210264 *) indexing queue slots are now configurable via config file
19 years ago
theli 3c11d7b81c *) Bugfix for minimizeUrlDB
19 years ago
orbiter 9913049009 fixed outOfMemory bug caused by loops in kelondroTree during enumeration
19 years ago
theli bbb936b9ea *) Bugfix for not human readable content of PDFs while viewing the URL Content via GUI
19 years ago
theli 445e3a620f *) Avoid rejecting of html content by the crawler when the file extension is not set properly
19 years ago
theli 444a5a9368 *) Bugfix for Entries with null url in GlobalQueue
19 years ago
borg-0300 ebac51df52 restore defaultRemoteProfile
19 years ago
borg-0300 5778428455 move cutUrlText to nxTools,
20 years ago
borg-0300 9158845c3b bugfix for snippet text null bytes
20 years ago
orbiter f763923e0a added missing files for last commit
20 years ago
orbiter 79818a320f introduced citation-rank transmission protocol and activate transport for anonymisation
20 years ago
theli 7e0647f692 *) Bugfix for userDB usage during authentication
20 years ago
orbiter 02f8013013 auto-delete of corrupted word files during word-migration
20 years ago
orbiter d2731418bf added creation of global ranking files and changed url normal form usage
20 years ago
theli 6f9f8ed8f8 *) Automatic Reset of Stack Crawler DB on startup errors
20 years ago