Problem with using fopen

6 Abr. 2021

0 Respuestas

Actualizado a las 6 Abr. 2021

9 Visualizaciones (30 días)

Iniciar sesión para responder a esta pregunta.

Follow Question

Iniciar sesión para responder a esta pregunta.

Follow Question

Mostrar comentarios más antiguos

0 votos

TestCOA.pdf

The goal is not just get the words from a pdf like you get from extractFileText(filename) syntax, but also the position of each sentence. The solution i use is to read the pdf and then flatedecode it to acive this information. After decoding the information can look like this:

I found a pyhonscript* that works and i want to translate it into matlab.

...here comes the problem

Python:

pdf = open("TestCOA.pdf","rb").read() <--- python read the file perfectly

Matlab:

fileID = fopen("TestCOA.pdf",'rb','n','us-ascii');

A = fscanf(fileID,'%c') <-- reads some char but mixed with invalid characters <?>

pdf=py.open("TestCOA.pdf","rb").read() <-- same results with the python integration syntax

Upploaded example pdf to try it out. Hope someone can help me to figure this out. :)

*The full python script: https://gist.github.com/averagesecurityguy/ba8d9ed3c59c1deffbd1390dafa5a3c2

0 comentarios
Mostrar -2 comentarios más antiguos Ocultar -2 comentarios más antiguos

Iniciar sesión para comentar.

Iniciar sesión para responder a esta pregunta.

Follow Question

Respuestas (0)

Iniciar sesión para responder a esta pregunta.

Categorías

Más información sobre Startup and Shutdown en Centro de ayuda y File Exchange.

Productos

MATLAB

Etiquetas

el 6 de Abr. de 2021

el 6 de Abr. de 2021

Community Treasure Hunt

Find the treasures in MATLAB Central and discover how the community can help you!

Translated by