Beautifulsoup - How to open images and download them

1 Answers

This will get you all URL of the images:

import urllib2
from bs4 import BeautifulSoup

url = "http://icecat.biz/p/toshiba/pscbxe-01t00een/satellite-pro-notebooks-4051528049077-Satellite+Pro+C8501GR-17732197.html"
html = urllib2.urlopen(url)
soup = BeautifulSoup(html)

imgs = soup.findAll("div", {"class":"thumb-pic"})
for img in imgs:
        print img.a['href'].split("imgurl=")[1]

Output:

http://www.toshiba.fr/contents/fr_FR/SERIES_DESCRIPTION/images/g1_satellite-pro-c850.jpg
http://www.toshiba.fr/contents/fr_FR/SERIES_DESCRIPTION/images/g4_satellite-pro-c850.jpg
http://www.toshiba.fr/contents/fr_FR/SERIES_DESCRIPTION/images/g2_satellite-pro-c850.jpg
http://www.toshiba.fr/contents/fr_FR/SERIES_DESCRIPTION/images/g5_satellite-pro-c850.jpg
http://www.toshiba.fr/contents/fr_FR/SERIES_DESCRIPTION/images/g3_satellite-pro-c850.jpg

And this code is for downloading and saving those images:

import os
import urllib
import urllib2
from bs4 import BeautifulSoup

url = "http://icecat.biz/p/toshiba/pscbxe-01t00een/satellite-pro-notebooks-4051528049077-Satellite+Pro+C8501GR-17732197.html"
html = urllib2.urlopen(url)
soup = BeautifulSoup(html)

imgs = soup.findAll("div", {"class":"thumb-pic"})
for img in imgs:
        imgUrl = img.a['href'].split("imgurl=")[1]
        urllib.urlretrieve(imgUrl, os.path.basename(imgUrl))

133

answered Nov 15 '22 15:11

4d4c

Related questions
                            
                                Extract external contour or silhouette of image in Python
                            
                                Paginate Django formset
                            
                                Checking a file existence on a remote SSH server using Python
                            
                                Can I open an application from a script during runtime?
                            
                                How can I get the android kernel version via adb (or via Python command)?
                            
                                Sending Mailgun Inline Images in HTML using Python Requests library
                            
                                Python: Fibonacci Sequence
                            
                                How to get all the hyponyms of a word/synset in python nltk and wordnet?
                            
                                Create Multidimensional Zeros Python
                            
                                IOError: [Errno 2] No such file - Paramiko put()
                            
                                multiprocess.apply_async How do I wrap *args and **kwargs?
                            
                                Delete related object via OneToOneField
                            
                                regex to get all text outside of brackets
                            
                                Python list([]) and []
                            
                                How to list only down-most directories in Python?
                            
                                How can one customize Django Rest Framework serializers output?
                            
                                python sort list based on key sorted list
                            
                                How to install PyQt5 on a new virtualenv and work on an IDLE
                            
                                python: hybrid between regular method and classmethod
                            
                                Deploy flask application on 1&1 shared hosting (with CGI)

Donate For Us

If you love us? You can donate to us via Paypal or buy me a coffee so we can maintain and grow! Thank you!

Donate Us With

Beautifulsoup - How to open images and download them

Tags:

python

beautifulsoup

Ninja2k

People also ask

1 Answers

4d4c

Recent Activity

Donate For Us