Dr. Mitchell Swartz wrote:

 For those interested, I have been working with Dr. Brian Josephson
many months, as an experiment, sending some of the papers by
email to Jed, cc to Brian.

  The fact is: Brian and the others received them.
 Brian encouraged Jed to put them up on his site, but Jed insisted
he edit them. Thereafter, Jed always found a "problem".

I insisted they be converted to text Acrobat format. I did not propose to change the content at all. Here are some of the letters from that exchange. I normally do not disclose this sort of thing, but I shall make an exception.

As I said here the other day, authors who do not wish to provide the papers in text Acrobat format are invited to take a hike. Anyone who finds this requirement onerous or unfair is invited to run your own damn web site. Yesterday I learned that Swartz has uploaded some of his own papers to his own site in HTML text format, for crying out loud. Why should he object to text Acrobat format?!? This makes zero sense.

So far, with ~1000 papers uploaded, only one author has objected to the format and organization of LENR-CANR.org: Swartz.

By the way, converting from HTML text to Acrobat text takes exactly as much time as it takes to print the document. I have a print driver that redirects the output to an Acrobat file. It's a piece of cake.

As readers here know, lately I have bent the standards to allow image-over-text format, because I got books in that format hundreds of pages long, such as the NSF/EPRI proceedings. This is not the same as image Acrobat format. I spent about a month correcting the underlying text in those books. There are still errors, unfortunately. (The text is "underlying" or "hidden" you might say. You can do a Control-F to find it, and Google finds it.)

- Jed

- - - - - - - - - - - - - -

FROM Swartz, 11 Jun 2006

Brian, and Jed:


Previously, I wrote:

"Our papers could be available in pdf image format, as I have stated before
if:  1) we are satisfied with the image quality,
2) we have the opportunity to check each prior to them being made available, and
     3) there is no editing of the paper after we ok the pdf image.
Previously, we have given pdf image papers to Jed of several papers, but he was wanted to use
his OCR on them, and that is not acceptable at this point in time.
We will allow OCR of the abstracts which can then be re-added in a non-image format
attached to the pdf-image papers,  as long as 1, 2, and 3 above are used, and
we have opportunity to check the OCR'd abstract text.
   In my experience, this use of pdf imaging creates a documents which is
a few hundred Kbytes to, at most, 2 megabytes if there are lots of graphic images and photographs. That should take care of all the problems cited in the missive below.
  Hope that helps.  Let me know if this is acceptable."

If this is acceptable, I will ask Alan to scan the CAM paper into pdf-image and port the pdf-image paper to Jed. I will be meeting with him shortly to go over some
experiments now ongoing here.  Let us know.

   Mitchell

- - - - - - - - - - - - - -

FROM ME TO Swartz and Josephson 11 Jun 2006

Mitchell Swartz writes:

>"Our papers could be available in pdf image format, as I have stated before
>if:  1) we are satisfied with the image quality,

Image format is not good. It does not index properly so people cannot find it; it is too big, and people have trouble reading it. It has to be Acrobat text format. Only the figures and equations should be images.

Let me know if you want my assistance converting to text. My OCR program is the best around.

- Jed

- - - - - - - - - - - - - -

FROM ME TO Swartz and Josephson 12 Jun 2006

Let me make it clear that I must insist on the text Acrobat format. That is our standard, and it works well for both the readers and for the mechanics of the web site. We have had great success distributing papers smoothly and rapidly all over the world, at the rate of 20,000 to 30,000 per month lately. I am confident that I know what I am doing, and I know what presentation format is called for.

Making the conversion is only a minor imposition on the author. So far no author has complained. If you do not have time to make the conversion now, please contact me when you have some free time and you are ready to upload the paper. Let me know if you would like assistance in the initial OCR phase. It should take you a day or two per paper to prepare. If I were you, I would re-enter the equations from scratch, and insert the best available graphs and other images. But that is up to you. As long as the base format is Acrobat text, the document will work with our system.

Of course the content and formatting of the document is entirely up to you.

- Jed

- - - - - - - - - - - - - -

--On 12 June 2006 00:25:39 -0400 Jed Rothwell <[email protected]> wrote:

As long as the base format is Acrobat text, the document will work
with our system.

Fine! So start with the abstract in that format, and include the rest of the paper, as scanned, as if it were graphics. Problem solved! Next?

Brian

- - - - - - - - - - - - - -

[I told Brian, no, sorry. No exceptions for any author.]

- - - - - - - - - - - - - -

FROM ME TO Swartz and his coauthors 13 June 2006

. . . Brian Josephson asked me to enumerate the main reasons I insist on the text Acrobat format. Here is what I told him:

1. It looks more professional and neat. Like it or not, thousands of readers judge the quality of a paper or web site by neatness. They will not even read a paper that looks messy. Since we know our readers tend to be that way, we must strive to make a good impression on them.

2. Image files are difficult for many people to read, and large, and unwieldy. Many of our readers in Russia, China, the Middle East and elsewhere must use slow connections. Scientists in Iran and Russia have such difficulty, they sometimes ask before a CD-ROM copy of the system (which I am happy to provide).

3. Google and the other search tools will not properly index an image Acrobat paper. The hybrid image-text format is even worse. Most of our readers come via Google, Yahoo and the other search tools, so we must make the papers visible to them, and correctly indexed.

4. They are compatible with tools used by disabled people, such as voice output, vision enhancement, Braille output, and special cursor controls.

5. Text Acrobat files are compatible with electronic dictionaries and translation tools. Many of our readers are outside of the US, and many are not native speakers of English so I expect they use electronic dictionaries to look up words, or they copy words or phrases to Google. I often do this with papers in Japanese.

6. Text format makes it much easier to look up references and quote text.

7. Text format allows important electronic content enhancements such as hyperlinks, contextual information, rejuvenation and so on. See: http://arxiv.org/help/faq/whytex

I think these advantages far outweigh the minor imposition on the authors, who must spend a few hours converting back into a machine readable format. Of course it is best to preserve the original machine readable files in the first place. If you happen to have the files for these papers, I can probably read them, even if the format is obsolete.

- Jed

Reply via email to