Hi,

recently we've migrated our ETD repository to DSpace which was created 
in 2001. We've also migrated the apache_logs to solr, but althought 
we've refused a lot of logs which were detected as robots, there are 
still a lot of records: 514millions! ...the solr index is about 50Gb....

Tomcat couldn't open and read this huge index because the Java heap 
space, so we resized the ram of the host to 8gb and finally it could 
open it...but now the statistics are so slow. I've been searching 
through solr and lucene posts to optimize all of this without success.

Someone has got a huge solr index or has some idea to improve this 
behaviour?

Thanks in advance,

Regards,

Jesús

-- 
.......................................................................
       __
     /   /       Jesús Martín García
C E / S / C A   Tècnic de Projectes
   /__ /         Centre de Serveis Científics i Acadèmics de Catalunya

Gran Capità, 2-4 (Edifici Nexus) · 08034 Barcelona
T. 93 551 6213 · F. 93 205 6979 · [email protected]
.......................................................................


------------------------------------------------------------------------------
All the data continuously generated in your IT infrastructure contains a
definitive record of customers, application performance, security
threats, fraudulent activity and more. Splunk takes this data and makes
sense of it. Business sense. IT sense. Common sense.
http://p.sf.net/sfu/splunk-d2dcopy1
_______________________________________________
DSpace-tech mailing list
[email protected]
https://lists.sourceforge.net/lists/listinfo/dspace-tech

Reply via email to