背景:
1、手上有几个大的xml文件,基本都在300M至600M之间;
2、XML内容包括title,co-author,abstract,Affiliation等;
3、用的是xmlreader进行解析;
遇到的问题:
如果解析所有内容,经常只能把XML文件的一部分解析出来,似乎是内存不够的迹象;
如果只把title或Affiliation单独解析出来,就能全部解析XML文件;
附上代码:
set_time_limit(0);
header("Content-Type: text/html;charset=utf-8");
$num=0;
$reader = new XMLReader();
$reader->open("JACS.xml");
while ($reader->read()) {
if($reader->nodeType==XMLREADER::ELEMENT) {
if ($reader->localName == "PubmedArticle") {
$num++;
echo 'Number:'.$num;
while ($reader->read()) {
if ($reader->nodeType == XMLREADER::ELEMENT) {
if ($reader->localName == "PubDate") {
while ($reader->read()){
if ($reader->nodeType == XMLREADER::ELEMENT) {
if ($reader->localName == "Year") {
$reader->read();
echo 'PublicationDate:'.$reader->value.' ';
break;
}
}
}
while ($reader->read()){
if ($reader->nodeType == XMLREADER::ELEMENT) {
if ($reader->localName == "Month") {
$reader->read();
echo $reader->value.' ';
break;
}
}
}
while ($reader->read()){
if ($reader->nodeType == XMLREADER::ELEMENT) {
if ($reader->localName == "Day") {
$reader->read();
echo $reader->value;
break;
}
}
}
echo '
';
break;
}
}
}
while ($reader->read()) {
if ($reader->nodeType == XMLREADER::ELEMENT) {
if ($reader->localName == "Title") {
$reader->read();
echo 'JournalName:'.$reader->value.'
';
break;
}
}
}
while ($reader->read()) {
if ($reader->nodeType == XMLREADER::ELEMENT) {
if ($reader->localName == "ArticleTitle") {
$reader->read();
echo 'ArticleTitle:'.$reader->value.'
';
break;
}
}
}
while ($reader->read()) {
if ($reader->nodeType == XMLREADER::ELEMENT) {
if ($reader->localName == "AbstractText") {
$reader->read();
echo 'Abstract:'.$reader->value.'
';
break;
}
}
}
while ($reader->read()) {
if ($reader->nodeType == XMLREADER::ELEMENT) {
if ($reader->localName == "Affiliation") {
$reader->read();
echo 'Affiliation:'.$reader->value.'
';
break;
}
}
}
}
}
}
$reader->close();
}
回复内容:
背景:
1、手上有几个大的xml文件,基本都在300M至600M之间;
2、XML内容包括title,co-author,abstract,Affiliation等;
3、用的是xmlreader进行解析;
遇到的问题:
如果解析所有内容,经常只能把XML文件的一部分解析出来,似乎是内存不够的迹象;
如果只把title或Affiliation单独解析出来,就能全部解析XML文件;
附上代码:
set_time_limit(0);
header("Content-Type: text/html;charset=utf-8");
$num=0;
$reader = new XMLReader();
$reader->open("JACS.xml");
while ($reader->read()) {
if($reader->nodeType==XMLREADER::ELEMENT) {
if ($reader->localName == "PubmedArticle") {
$num++;
echo 'Number:'.$num;
while ($reader->read()) {
if ($reader->nodeType == XMLREADER::ELEMENT) {
if ($reader->localName == "PubDate") {
while ($reader->read()){
if ($reader->nodeType == XMLREADER::ELEMENT) {
if ($reader->localName == "Year") {
$reader->read();
echo 'PublicationDate:'.$reader->value.' ';
break;
}
}
}
while ($reader->read()){
if ($reader->nodeType == XMLREADER::ELEMENT) {
if ($reader->localName == "Month") {
$reader->read();
echo $reader->value.' ';
break;
}
}
}
while ($reader->read()){
if ($reader->nodeType == XMLREADER::ELEMENT) {
if ($reader->localName == "Day") {
$reader->read();
echo $reader->value;
break;
}
}
}
echo '
';
break;
}
}
}
while ($reader->read()) {
if ($reader->nodeType == XMLREADER::ELEMENT) {
if ($reader->localName == "Title") {
$reader->read();
echo 'JournalName:'.$reader->value.'
';
break;
}
}
}
while ($reader->read()) {
if ($reader->nodeType == XMLREADER::ELEMENT) {
if ($reader->localName == "ArticleTitle") {
$reader->read();
echo 'ArticleTitle:'.$reader->value.'
';
break;
}
}
}
while ($reader->read()) {
if ($reader->nodeType == XMLREADER::ELEMENT) {
if ($reader->localName == "AbstractText") {
$reader->read();
echo 'Abstract:'.$reader->value.'
';
break;
}
}
}
while ($reader->read()) {
if ($reader->nodeType == XMLREADER::ELEMENT) {
if ($reader->localName == "Affiliation") {
$reader->read();
echo 'Affiliation:'.$reader->value.'
';
break;
}
}
}
}
}
}
$reader->close();
}
可以参考一下 这个 PHP处理比较大的XML文件
GNU makefile中文手册 pdf,文比较完整的讲述GNU make工具,涵盖GNU make的用法、语法。同时重点讨论如何为一个工程编写Makefile。阅读本书之前,读者应该对GNU的工具链和Linux的一些常用编程工具有一定的了解。诸如:gcc、as、ar、ld、yacc等本文比较完整的讲述GNU make工具,涵盖GNU make的用法、语法。重点讨论如何使用make来管理软件工程、以及如何为工程编写正确的Makefile。 本手册不是一个纯粹的语言翻译版本,其中对GNU make的一些语法
为啥要装那么大 txt打开那么大也死机了 多分几个文件吧









