30行Python代码就可以调用ChatGPT API总结论文的主要内容- 汇维网

阅读论文可以说是我们的日常工作之一，论文的数量太多，我们如何快速阅读归纳呢？自从ChatGPT出现以后，有很多阅读论文的服务可以使用。其实使用ChatGPT API非常简单，我们只用30行python代码就可以在本地搭建一个自己的应用。

使用 Python 和 ChatGPT API 总结论文的步骤很简单：

用于 PDF 处理的 PyPDF2 和用于与 GPT-3.5-turbo 接口的 OpenAI。
使用 PyPDF2 打开并阅读 PDF 文件。
遍历 PDF 文档中的每一页，提取文本。
使用 GPT-3.5-turbo 为每个页面的文本生成摘要。
合并摘要并将最终摘要文本保存到文件中。

import PyPDF2
 import openai
 pdf_summary_text = ""

解析pdf

pdf_file_path = "./pdfs/paper.pdf"
 pdf_file = open(pdf_file_path, 'rb')
 pdf_reader = PyPDF2.PdfReader(pdf_file)

获取每一页的文本：

for page_num in range(len(pdf_reader.pages)):
    page_text = pdf_reader.pages[page_num].extract_text().lower()

使用openai的api进行汇总

response = openai.ChatCompletion.create(
    model="gpt-3.5-turbo",
    messages=[
        {"role": "system", "content": "You are a helpful research assistant."},
        {"role": "user", "content": f"Summarize this: {page_text}"},
    ],
 )
 page_summary = response["choices"][0]["message"]["content"]

合并摘要

pdf_summary_text += page_summary + "n"
 pdf_summary_file = pdf_file_path.replace(os.path.splitext(pdf_file_path)[1], "_summary.txt")
 with open(pdf_summary_file, "w+") as file:
    file.write(pdf_summary_text)

搞定，关闭pdf文件，回收内存

pdf_file.close()

完整代码如下：

import os
 import PyPDF2
 import re
 import openai
 
 # Here I assume you are on a Jupiter Notebook and download the paper directly from the URL
 !curl -o paper.pdf https://arxiv.org/pdf/2301.00810v3.pdf?utm_source=pocket_saves
 
 # Set the string that will contain the summary    
 pdf_summary_text = ""
 # Open the PDF file
 pdf_file_path = "paper.pdf"
 # Read the PDF file using PyPDF2
 pdf_file = open(pdf_file_path, 'rb')
 pdf_reader = PyPDF2.PdfReader(pdf_file)
 # Loop through all the pages in the PDF file
 for page_num in range(len(pdf_reader.pages)):
    # Extract the text from the page
    page_text = pdf_reader.pages[page_num].extract_text().lower()
     
    response = openai.ChatCompletion.create(
                    model="gpt-3.5-turbo",
                    messages=[
                        {"role": "system", "content": "You are a helpful research assistant."},
                        {"role": "user", "content": f"Summarize this: {page_text}"},
                            ],
                                )
    page_summary = response["choices"][0]["message"]["content"]
    pdf_summary_text+=page_summary + "n"
    pdf_summary_file = pdf_file_path.replace(os.path.splitext(pdf_file_path)[1], "_summary.txt")
    with open(pdf_summary_file, "w+") as file:
        file.write(pdf_summary_text)
 
 pdf_file.close()
 
 with open(pdf_summary_file, "r") as file:
    print(file.read())

需要说明的是2个事情：

1、openai的API免费调用额度是有限的，这个方法一篇论文大概在0.2-0.5美元左右，根据论文长度会有变化

2、gpt4的API我没测试，因为我还没有申请到，并且看价格那个太贵了（贵20倍）我觉得不值，但是可以试试把论文的图表一同传过去，是不是会有更好效果（不确定）

1 原创文章作者：汉墨堂-总部，如若转载，请注明出处： https://www.52hwl.com/63340.html

2 温馨提示：软件侵权请联系469472785#qq.com（三天内删除相关链接）资源失效请留言反馈

3 下载提示：如遇蓝奏云无法访问，请修改lanzous(把s修改成x)

4 免责声明：本站为个人博客，所有软件信息均来自网络修改版软件，加群广告提示为修改者自留，非本站信息，注意鉴别

30行Python代码就可以调用ChatGPT API总结论文的主要内容

关于作者

汉墨堂-总部注册用户

30行Python代码就可以调用ChatGPT API总结论文的主要内容

关于作者

汉墨堂-总部注册用户

相关推荐

微软必应再强化！接入OpenAI DALL·E模型，文字生成图像

清华大学周伯文：ChatGPT火爆揭示新一代协同与交互智能的高度重要性

人工智能导致可怕的网络安全威胁的三种方式

人工智能监管：向前迈进还是道德清洗?

ChatGPT一路狂飙，我们的安全带在哪儿？

扩散模型生成带汉字图像，一键输出表情包：OPPO等提出GlyphDraw