Q: 如何把jupyter notebook 转为 pdf 文档?

A: 尝试了几种python包, 结果都没有成功. 包括: xhtml2pdf,

查看官方的介绍说用pandoc也是一种方法, 但是觉得安装一个可怕的Latex和pandoc太麻烦了.

还好, 找到了一个开源方法: 用wkhtmltopdf 程序.

用python写一个脚本, 调用wkhtmltopdf, 运行命令行指令, 得以实现. 非常符合我的预期. 简明, 优雅.

wkhtml2pdf 简介

wkhtmltopdf，一个集成好了的exe文件（C++编写），

基本的调用方法是:

"c:\Program Files\bin\wkhtmltopdf.exe" https://github.com/mementum/backtrader/blob/master/docs2/signal_strategy/signal_strategy.rst signal_strategy.pdf

Loading pages (1/6)

Counting pages (2/6)

Resolving links (4/6)

Loading headers and footers (5/6)

Printing pages (6/6)

Done

C:\Documents and Settings\Administrator\duanqs\strategy_study>dir *.pdf

 驱动器 C 中的卷是 160GB_XP

 卷的序列号是 EC5F-C44B

 C:\Documents and Settings\Administrator\duanqs\strategy_study 的目录

2017-04-17  14:47           120,295 signal_strategy.pdf

2017-04-17  13:32           597,111 backtest.pdf

               2 个文件        717,406 字节

               0 个目录 19,999,031,296 可用字节

可以先在命令行测试一下，有其他的需要, 可以在命令行通过wkhtmltopdf --help查询，

如果是超长页的话，可以用命令:

wkhtmltopdf.exe http://passport.yupsky.com/ac/register e:\yupskyreg.pdf -H --outline

Here:

-H 是显示扩展帮助

--outline 是添加pdf的左侧概要！(缺省设置)

而且可以批量生成哦，中间用空格隔开

python 脚本: (封装了运行wkhtml2pdf.exe 命令行的py脚本)



# code:utf-8

'''

IPython/Jupyter Problems saving notebook as PDF - Stack Overflow

http://stackoverflow.com/questions/29156653/ipython-jupyter-problems-saving-notebook-as-pdf

This Python script has GUI to select with explorer a Ipython Notebook you want to convert to pdf.

The approach with wkhtmltopdf is the only approach I found works well and provides high quality pdfs.

Other approaches described here are problematic, syntax highlighting does not work or graphs are messed up.

You'll need to install wkhtmltopdf: http://wkhtmltopdf.org/downloads.html

and Nbconvert

pip install nbconvert

# OR

conda install nbconvert

'''

# Script adapted from CloudCray

# Original Source: https://gist.github.com/CloudCray/994dd361dece0463f64a

# 2016--06-29

# This will create both an HTML and a PDF file

import subprocess

import os

from Tkinter import Tk

from tkFileDialog import askopenfilename

WKHTMLTOPDF_PATH = "C:/Program Files/wkhtmltopdf/bin/wkhtmltopdf"  # or wherever you keep it

def export_to_html(filename):

    cmd = 'ipython nbconvert --to html "{0}"'

    subprocess.call(cmd.format(filename), shell=True)

    return filename.replace(".ipynb", ".html")

def convert_to_pdf(filename):

    cmd = '"{0}" "{1}" "{2}"'.format(WKHTMLTOPDF_PATH, filename, filename.replace(".html", ".pdf"))

    subprocess.call(cmd, shell=True)

    return filename.replace(".html", ".pdf")

def export_to_pdf(filename):

    fn = export_to_html(filename)

    return convert_to_pdf(fn)

def main():

    print("Export IPython notebook to PDF")

    print("    Please select a notebook:")

    Tk().withdraw() # Starts in folder from which it is started, keep the root window from appearing

    x = askopenfilename() # show an "Open" dialog box and return the path to the selected file

    x = str(x.split("/")[-1])

    print(x)

    if not x:

        print("No notebook selected.")

        return 0

    else:

        fn = export_to_pdf(x)

        print("File exported as:\n\t{0}".format(fn))

        return 1

main()

这里也记录一下尝试xhtml2pdf的经过.

安装完了以后, 编写脚本, 运行时主要是: 卡在了html5lib这个包里:

异常是:

inputstream

CSS parser

等等.

搞定不了, 所以放弃之.

install xhtml2pdf and update html5lib from old vertion to new version (1.0b8)

Here is the logging:

C:\Documents and Settings\Administrator>pip install xhtml2pdf

Collecting xhtml2pdf

  Downloading xhtml2pdf-0.0.6.zip (120kB)

    100% |████████████████████████████████| 122kB 467kB/s

Collecting html5lib (from xhtml2pdf)

  Using cached html5lib-0.999999999-py2.py3-none-any.whl

Collecting pyPdf2 (from xhtml2pdf)

  Downloading PyPDF2-1.26.0.tar.gz (77kB)

    100% |████████████████████████████████| 81kB 10kB/s

Requirement already satisfied: Pillow in d:\anaconda2\lib\site-packages (from xhtml2pdf)

Collecting reportlab>=2.2 (from xhtml2pdf)

  Downloading reportlab-3.4.0-cp27-cp27m-win32.whl (2.1MB)

    100% |████████████████████████████████| 2.1MB 261kB/s

Collecting webencodings (from html5lib->xhtml2pdf)

  Downloading webencodings-0.5.1-py2.py3-none-any.whl

Requirement already satisfied: setuptools>=18.5 in d:\anaconda2\lib\site-packages (from html5lib->xhtml2pdf)

Requirement already satisfied: six in d:\anaconda2\lib\site-packages (from html5lib->xhtml2pdf)

Requirement already satisfied: pip>=1.4.1 in d:\anaconda2\lib\site-packages (from reportlab>=2.2->xhtml2pdf)

Requirement already satisfied: packaging>=16.8 in d:\anaconda2\lib\site-packages (from setuptools>=18.5->html5lib->xhtml

2pdf)

Requirement already satisfied: appdirs>=1.4.0 in d:\anaconda2\lib\site-packages (from setuptools>=18.5->html5lib->xhtml2

pdf)

Requirement already satisfied: pyparsing in d:\anaconda2\lib\site-packages (from packaging>=16.8->setuptools>=18.5->html

5lib->xhtml2pdf)

Building wheels for collected packages: xhtml2pdf, pyPdf2

  Running setup.py bdist_wheel for xhtml2pdf ... done

  Stored in directory: C:\Documents and Settings\Administrator\Local Settings\Application Data\pip\Cache\wheels\ec\eb\db

\13a2be9c15f492c65086709a69042924ebfb7aa4c4cc7284f1

  Running setup.py bdist_wheel for pyPdf2 ... done

  Stored in directory: C:\Documents and Settings\Administrator\Local Settings\Application Data\pip\Cache\wheels\86\6a\6a

\1ce004a5996894d33d93e1fb1b67c30973dc945cc5875a1dd0

Successfully built xhtml2pdf pyPdf2

Installing collected packages: webencodings, html5lib, pyPdf2, reportlab, xhtml2pdf

Successfully installed html5lib-0.999999999 pyPdf2-1.26.0 reportlab-3.4.0 webencodings-0.5.1 xhtml2pdf-0.0.6

C:\Documents and Settings\Administrator>pip install html5lib==1.0b8

Collecting html5lib==1.0b8

  Downloading html5lib-1.0b8.tar.gz (889kB)

    100% |████████████████████████████████| 890kB 311kB/s

Requirement already satisfied: six in d:\anaconda2\lib\site-packages (from html5lib==1.0b8)

Building wheels for collected packages: html5lib

  Running setup.py bdist_wheel for html5lib ... done

  Stored in directory: C:\Documents and Settings\Administrator\Local Settings\Application Data\pip\Cache\wheels\d4\d1\0b

\a6b6f9f204af55c9bb8c97eae2a78b690b7150a7b850bb9403

Successfully built html5lib

Installing collected packages: html5lib

  Found existing installation: html5lib 0.999999999

    Uninstalling html5lib-0.999999999:

      Successfully uninstalled html5lib-0.999999999

Successfully installed html5lib-1.0b8

C:\Documents and Settings\Administrator>

ipynb to pdf的更多相关文章

Windows7下Jupyter Notebook使用入门
目录一.Jupyter简介二.Jupyter安装 2.1 python 3安装 2.2 Jupyter 安装三.Jupyter使用示例四.Jupyter常用命令五.其他说明一.Jupyte ...
简单python脚本，将jupter notebook的ipynb文件转为pdf（包含中文）
直接执行的python代码ipynb2pdf.py 主要思路.将ipynb文件转成tex文件,然后使用latex编译成pdf.由于latex默认转换不显示中文,需要向tex文件中添加相关中文包. 依赖 ...
windows jupyter lab中.ipynb转中文PDF
在jupyter lab中,File-Export Notebook as-Export Notebook to PDF,可以导出成PDF格式的文档,但在操作前需要安装些程序.1. 安装pandocA ...
Jupyter Notebook PDF输出的中文支持
Jupyter Notebook是什么 Jupyter Notebook是ipython Notebook 的升级.Jupyter能够将实时代码,公式,可视化图表以Cell的方式组织在一起,形成一个对 ...
Jupyter Notebook通过latex输出pdf
主要步骤 1.将ipynb编译成tex ipython nbconvert --to latex Example.ipynb 2. 修改tex,增加中文支持在\documentclass{artic ...
是程序员，就用python导出pdf
这两天一直在做课件,我个人一直不太喜欢PPT这个东西--能不用就不用,我个人特别崇尚极简风. 谁让我们是程序员呢,所以就爱上了Jupyter写课件,讲道理markdown也是个非常不错的写书格式啊. ...
Python学习笔记——jupyter notebook 入门和中文pdf输出方案
简单粗暴的安装对于懒人而言,我还是喜欢直接安装python的集成开发环境 anaconda 多个内核控制 jupyter官网 1). 同时支持python2 和python 3 conda crea ...
【原创】JavaFx程序解决Jupyter Notebook导出PDF不显示中文
0.ATTENTION!!! JavaFx里是通过Java调用控制台执行的的jupyter和xelatex指令, 这些个指令需要在本地安装Jupyter和MikTeX之后才能正常在电脑上运行 1.[问 ...
C#给PDF文档添加文本和图片页眉
页眉常用于显示文档的附加信息,我们可以在页眉中插入文本或者图形,例如,页码.日期.公司徽标.文档标题.文件名或作者名等等.那么我们如何以编程的方式添加页眉呢?今天,这篇文章向大家分享如何使用了免费组件 ...

随机推荐

如何实现MySQL表数据随机读取?从mysql表中读取随机数据
文章转自 http://blog.efbase.org/2006/10/16/244/如何实现MySQL表数据随机读取?从mysql表中读取随机数据?以前在群里讨论过这个问题,比较的有意思.mysql ...
2013成都网赛1010 hdu 4737 A Bit Fun
题意:定义f(i, j) = ai|ai+1|ai+2| ... | aj (| 指或运算),求有多少对f(i,j)<m.1 <= n <= 100000, 1 <= m &l ...
Dubbo学习(一) Dubbo原理浅析
一.初入Dubbo Dubbo学习文档: http://dubbo.incubator.apache.org/books/dubbo-user-book/ http://dubbo.incubator ...
mock测试SpringMVC controller报错
使用mock测试Controller时报错如下 java.lang.NoClassDefFoundError: javax/servlet/SessionCookieConfig at org.spr ...
对synchronized的一点理解
一.synchronized的使用(一).synchronized同步方法1. “非线程安全”问题存在于“实例变量”中,如果是方法内部的私有变量,则不存在“非线程安全”问题.2. 如果多个线程共同访问 ...
03.基于IDEA+Spring+Maven搭建测试项目--常用dependency
  <properties> <java.version>1.8</j ...
Stone Game, Why are you always there? HDU - 2999（sg定理）
题意:给你n个数的集合,表示你每次取石子只能为集合里的数,然后给你一排石子,编号为1~n,每次你可以取相邻位置的连续石子(数量只能为集合里的数),注意石子的位置时不变的,比如把2拿走了,1和3还是不相 ...
Integer to Roman - LeetCode
目录题目链接注意点解法小结题目链接 Integer to Roman - LeetCode 注意点考虑输入为0的情况解法解法一:从大到小考虑1000,900,500,400,100,9 ...
洛谷 P2473 [SCOI2008]奖励关解题报告
P2473 [SCOI2008]奖励关题目描述你正在玩你最喜欢的电子游戏,并且刚刚进入一个奖励关.在这个奖励关里,系统将依次随机抛出\(k\)次宝物,每次你都可以选择吃或者不吃(必须在抛出下一个宝 ...
第三周构造一个简单的Linux系统
20135331文艺首先在上周内容中我们学习了计算机三个法宝: 1.存储程序计算机 2.函数调用堆栈 3.中断本周中得知操作系统两把宝剑: 1.中断上下文的切换:保存现场和恢复现场 2.进程 ...

ipynb to pdf