Hi,
recently we've migrated our ETD repository to DSpace which was created
in 2001. We've also migrated the apache_logs to solr, but althought
we've refused a lot of logs which were detected as robots, there are
still a lot of records: 514millions! ...the solr index is about 50Gb....
Tomcat couldn't open and read this huge index because the Java heap
space, so we resized the ram of the host to 8gb and finally it could
open it...but now the statistics are so slow. I've been searching
through solr and lucene posts to optimize all of this without success.
Someone has got a huge solr index or has some idea to improve this
behaviour?
Thanks in advance,
Regards,
Jesús
--
.......................................................................
__
/ / Jesús Martín García
C E / S / C A Tècnic de Projectes
/__ / Centre de Serveis Científics i Acadèmics de Catalunya
Gran Capità, 2-4 (Edifici Nexus) · 08034 Barcelona
T. 93 551 6213 · F. 93 205 6979 · [email protected]
.......................................................................
------------------------------------------------------------------------------
All the data continuously generated in your IT infrastructure contains a
definitive record of customers, application performance, security
threats, fraudulent activity and more. Splunk takes this data and makes
sense of it. Business sense. IT sense. Common sense.
http://p.sf.net/sfu/splunk-d2dcopy1
_______________________________________________
DSpace-tech mailing list
[email protected]
https://lists.sourceforge.net/lists/listinfo/dspace-tech