orbiter
1dce2f1079
more multithreading support:
...
- replaced some synchronized classes by classes from util.concurrent
- used a util.concurrent.SynchronousQueue to implement a persistent sorting thread in
the very basic kelondroRowCollection which supports sorting with a second thread
in case that a double-core processing CPU is used
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4517 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
6779b455d7
another fix for the punycode parser/generator (should work now!)
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4516 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
1b127406d0
update to punycode encoding (still not working)
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4515 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
83860507c9
- added punycode class from gnu idn library
...
- added parser for international domains in yacyURL
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4514 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
253a453413
removed possible synchronization deadlock
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4511 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
3f321ece7d
added a search history to the new search page
...
the history distinguishes between different users and identifies them by their ip
a history is only shown to the user who submitted the search
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4510 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
c48e25d784
- fixed selection box for topwords
...
- fixed parser detail in condenser
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4509 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
87a8747ce3
- enhanced recognition, parsing, management and double-occurrence-handling of image tags
...
- enhanced text parser (condenser): found and eliminated bad code parts; increase of speed
- added handling of image preview using the image cache from HTCACHE
- some other minor changes
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4507 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
low012
652086159a
*) Replaced System.err.println() by logging function. Left System.err.println()s as comments to be able to quickly revert changes since gzip is an application with it's own main method and Orbiter maybe wants to keep it this way.
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4505 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
677ee2ea04
added remove operation to collection index (re-activation)
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4503 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
d477483373
stronger criteria to use RAM copy to use table copy
...
(should use less RAM)
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4502 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
a7abee3578
- fixed some data types in new search stack
...
- added image domain presentation to image preview
- added new search page to menu
- added automatic re-search when an old search profile is requested and a crawl is ongoing,
to fetch newly crawled entries
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4501 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
81687b6bd5
added missing hachCode computation for previous feature
...
this solves also the missing image double-check fetaure!
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4500 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
bedd8dfbe2
- added image sorting by image size. This is the default now.
...
This is performed using a 3-stage sorting process:
- sort by relevance, then do snippet-fetch
- sort snippets by relevance then do image link extraction
- sort image links by image size; unknown sizes are handled like small sizes
- only the exact amount of images as requested are shown
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4499 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
727feb4358
- fixed some bugs in ranking computation
...
- introduced generalized method to organize ranked results (2 new classes)
- added a post-ranking after snippet-fetch (before: only listed) using the new ranking data structures
- fixed some missing data fields in RWI ranking attributes and correct hand-over between data structures
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4498 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
f4c73d8c68
- fixed highslide usage
...
- some enhancement to index management, better types
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4497 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
2327451653
- changed order of database initialisation (index first)
...
- removed mainly unused init-time for databases (was only used for tree tables, which are not used any more)
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4496 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
3441ec3928
- some small changes to highslide integration to get it working... (does not work yet)
...
- performance enhancement for url list parser
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4495 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
6c3cd2b4f2
- added new way to watch images from the image search:
...
they appear as separate, floating window above the search results,
not in a new window
- added highslide javascript library for feature mentioned above
- removed dir servlet. This thing was not used as it was supposed to be (as an example applet)
and was a major problem for intranet-indexing when files are hosted on the same peer.
- added yacy-httpd-internal directory listing. Because YaCy is a search engine,
directory listings are similar to search result listings. Intranet indexing from the same peer
will get nice index pages for document collections.
- removed unused test applet
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4494 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
61a81820e3
- refactoring of search tracker
...
- added link to search history to repeat the search
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4493 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
lulabad
9ecc17baef
fixed double Blog entrys
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4492 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
36b898ca7a
- tested successfully z-presentation of yacy seed encoding
...
- added alternative switch that takes shortest representation as yacy seed string encoding
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4491 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
066c88140f
quickfix for OOM, see http://forum.yacy-websuche.de/viewtopic.php?f=5&t=875&hilit=&p=5686#p5686
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4488 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
4079c38ce0
- probably slightly better default ranking
...
- added experimental right column to new search page (no function, only container)
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4487 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
8fd5e52f04
added basket icons and experimental gif animation class
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4485 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
lulabad
94e256e13b
* removed single Blogview, now links direct to BlogComments.html
...
* some other small changes
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4483 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
ff5969901c
modified dir servlet to cooperate with intranet indexing from the own HTDOCS repository:
...
- removed md5 file generation (spoils the won repository)
- removed comments in file share (was never used)
- moved dir list comparator to other place (maybe solves problem, lets see)
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4481 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
lulabad
00f5f917de
- more refactoring to blog
...
- fixed moderate comment bug. see http://forum.yacy-websuche.de/viewtopic.php?f=9&t=860
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4478 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
f890b039ee
experiments wit openstreetmaps
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4477 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
7f445f34a6
bitte die Java 5 - typischen Warnings einschalten!
...
(unboxed-Fehler wies auf Programmfehler hin und Typangabe fehlte)
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4476 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
lulabad
c1b9a03304
* some refactoring to Blog
...
* changed default sort order to reverse (newest first)
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4475 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
lulabad
766a04bc06
fixed sort problem in Blog. see http://forum.yacy-websuche.de/viewtopic.php?f=6&t=639
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4474 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
borg-0300
bfe171e693
Small change (generics)
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4473 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
borg-0300
2589290ded
better ping
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4472 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
borg-0300
dae9053b21
BUGFIX
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4464 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
borg-0300
77ba446332
seedDB helpers update/cleanup
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4461 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
borg-0300
dd215e7f6b
NPE fix
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4460 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
bd63999801
- faster search: using different data structures that avoid multiplr calculations
...
- no more table copy for error-eco table
- optional table copy for lurl-entries
- more abstractions (less single constant strings)
- better logging (using host names instead of ips)
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4459 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
lulabad
8358652fa9
some small changes to blog
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4457 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
159aaf8889
re-introduced global search limitation when index receive is switched off
...
this was necessary because othervise robinson peers did also global searches, which cannot be a wanted effect
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4456 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
borg-0300
a9c4e9c309
Small change (ping)
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4453 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
borg-0300
9ab6ad8b73
more seedDB helpers
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4452 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
lulabad
6a85764e1a
Second bugfix for numberbug in Blog.
...
This update fix automatic existing blogentrys.
A backup is not needed but almost a good idea ;)
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4451 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
efd5807a7c
- some renaming of variables to support DC
...
- initial 120mb RAM for fresh peers
- release 0.57
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4445 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
lulabad
40a0591942
Fixed numberbug in Blog, see http://forum.yacy-websuche.de/viewtopic.php?f=6&t=639 . This wont fix existing Blogentrys (comes later).
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4443 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
141db7ba48
there is less RAM needed for eco table (its just a security-plus for RAM check)
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4442 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
249d61759a
fix for false RAM table activation in EcoTables
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4441 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
ff6b69b37e
fix for NPE in access tracker
...
fix for NPE in word index
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4439 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
3c7b94c119
- fix for online caution delay settings, see
...
http://forum.yacy-websuche.de/viewtopic.php?f=6&t=738&p=4723#p4723
- removed remote search limitation for non-dht-peers according to discussion in
http://forum.yacy-websuche.de/viewtopic.php?f=15&t=793&hilit=&p=5277#p5277
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4438 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
f35a3794e0
auto-healing (deletion) of bad peer addresses during start-up
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4437 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
42c1e11f2b
added another link double-check
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4434 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
a5d388bfff
fix for HTCache organisation that may have caused unlimited grow of the cache
...
appeared only for tree-caches
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4433 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
96c5e6acc7
added a double-check for search results
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4432 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
a1e9e6e2e6
fix for search result page navigation
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4431 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
7404256997
- no more search time-out!
...
- fixed a bug with last commit
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4430 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
cd3e0d6f03
tried to fix another eco bug
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4429 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
08a12e9bb5
- removed dashed line from default skin (looks much better!)
...
- better timing when displaying results
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4428 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
89169d54fd
fixed search result preparation
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4427 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
acf771d5e1
- fixed bug with too much RAM in crawler queue
...
- fixed dir bug
- better calculation of TF for join
- better waiting-on-result logic
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4424 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
a8a5df4a51
- more dublin core naming of page metadata
...
- better presentation of result counters in search results
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4420 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
fa3b8f0ae1
fixed bug in remote search
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4419 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
7d875290b2
more generics
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4417 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
9d693ee635
more generics
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4415 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
0f5c4abaca
more generics
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4414 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
974fea7933
added term-frequency ranking
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4413 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
1a296af6ff
more generics
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4412 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
4a80902081
- added ViewProfile as rdf in foaf syntax
...
- added link to rdf and vCard version on html page
- can be seen on http://localhost:8080/ViewProfile.html?hash=localhash
- more generics
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4411 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
da8c850a25
disabled IO path optimization (seems to block other methods)
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4405 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
hermens
d177ceb3b3
Fix for growing responseHeader[12].db when using proxyCacheLayout = hash
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4404 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
apfelmaennchen
b1fae9b5af
fixed import Netscape Bookmarks
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4401 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
2485681002
added termination control for RotateIterator
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4399 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
e2e7f065e9
minor fixes, some generics
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4398 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
15397298dc
- refactoring of indexControlRWIs: moved statics to own class; better Dublin Core naming
...
- fix for http://forum.yacy-websuche.de/viewtopic.php?f=5&t=759&hilit=&p=4866#p4866
- some bugfixes in EcoTable according remove method
- switched more tables to Eco: crawl Profiles, htcache, seeddb, newsdb
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4397 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
apfelmaennchen
f3a9e9c542
added getFolderList() to bookmarksDB
...
added cleanTagsString() to bookmarksDB
added getFoldersString() to Bookmark
modified getTagsString() to exclude folderTags
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4383 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
db25425893
more generics
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4382 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
9e7cd4fdbb
more generics
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4380 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
4e70dff8cf
more generics
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4379 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
6dc679785f
- fixed bad sort behavior of kelondroRowSet, in this case: no sort at all!
...
see http://forum.yacy-websuche.de/viewtopic.php?p=4841#p4841
- some memory calculation enhancements in kelondroFlex and a little bit more logging
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4378 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
0b4205eb5a
- fix double-deletion in eco tables
...
- changed behaviour of sort moment (not during a get)
- added some asserts in snippet cache for debugging
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4375 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
4ce6fab428
added special handling for doubles in eco tables after initialization
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4370 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
002a109c4d
patch for http://forum.yacy-websuche.de/viewtopic.php?p=4597#p4597
...
(urls that have no protocol but start with www will be treated as http://www ...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4369 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
634430c48a
- more logging
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4368 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
d372a78aef
some fixes to bring back lulabads peer..
...
see also: http://forum.yacy-websuche.de/viewtopic.php?p=4772#p4772
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4366 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
low012
f4799c2334
*) removed since I decided to turn this into a project of it's own using Perl to gather n-gram data which YaCy will be able to use
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4365 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
4ffbcd54a4
fix for http://forum.yacy-websuche.de/viewtopic.php?f=6&t=754
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4358 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
apfelmaennchen
e81bced2bd
reorganized the code and adjusted getTagIterator() to suit folders
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4357 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
85dc62c16f
refactoring: more dublin core - compliant naming
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4354 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
efd0b8371a
- added parsing of Dublin Core - compliant metadata (see RFC 5013 and ISO 15836) to html parser
...
- refactoring of plasmaParserDocument to use Dublin Core - compatible property names
- redesign of url handling in parser and condenser (less String-to-yacyURL conversion)
- more generics
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4352 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
low012
cfd4fecd12
*) blanks in paths for restart and update script are replaced by backslash+blank now (see http://forum.yacy-websuche.de/viewtopic.php?t=745 )
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4351 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
f945ee21d2
some security additions, keep maximum byte[] size to 2^27
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4350 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
2f3b2f3481
- extended dbtest for comparisment tests
...
- added initial space option for eco tables
- used initial space value in initialization of collectionIndex, this should avoid OOM failures" /Volumes/Magneto/dev/workspace/trunk/source/dbtest.java /Volumes/Magneto/dev/workspace/trunk/source/de/anomic/kelondro/kelondroCollectionIndex.java /Volumes/Magneto/dev/workspace/trunk/source/de/anomic/kelondro/kelondroDyn.java /Volumes/Magneto/dev/workspace/trunk/source/de/anomic/kelondro/kelondroEcoTable.java /Volumes/Magneto/dev/workspace/trunk/source/de/anomic/kelondro/kelondroRow.java /Volumes/Magneto/dev/workspace/trunk/source/de/anomic/kelondro/kelondroSplitTable.java /Volumes/Magneto/dev/workspace/trunk/source/de/anomic/plasma/plasmaCrawlBalancer.java /Volumes/Magneto/dev/workspace/trunk/source/de/anomic/plasma/plasmaCrawlStacker.java /Volumes/Magneto/dev/workspace/trunk/source/de/anomic/plasma/plasmaCrawlZURL.java
- added index consistency check (checks for double-occurrences of primary keys in file)
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4349 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
9eb746863d
interface enhancements for eco records memory statistics
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4348 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
9abc927645
to fix inconsistencies in collection index, a double reference reporting mechanism has been implemented
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4347 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
58a1f518f8
fixed some problems with eco tables
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4346 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
d4d07802ac
better RAM protection using eco tables
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4345 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
f4e9ff6ce9
more generics
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4343 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
cbefc651ac
more generics
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4342 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
45339c3db5
more generics
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4341 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
94f21d9403
activated new kelondroEcoTable file structure.
...
This data structure replaces almost all files in the PLASMA directory
also the collection.index and the LURL-db will be created as Eco-DB, if it does not exist before
existing Flex-databases will be used as they are (the is no data lost)
If you want to force the creation of a Eco-collection.index, simply delete the old index.
The Eco file system will only be used if there is enough memory.
The collection.index RAM limit is 200MB, if you have less, a flex-Table is createt.
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4340 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago
orbiter
a0f7f2faad
some more generics
...
git-svn-id: https://svn.berlios.de/svnroot/repos/yacy/trunk@4338 6c8d7289-2bf4-0310-a012-ef5d649a1542
17 years ago