Instant Jsoup How-to
eBook - ePub

Instant Jsoup How-to

Pete Houston

  1. 38 pages
  2. English
  3. ePUB (mobile friendly)
  4. Available on iOS & Android
eBook - ePub

Instant Jsoup How-to

Pete Houston

Book details
Book preview
Table of contents
Citations

About This Book

In Detail

As you might know, there are a lot of Java libraries that support parsing HTML content out there. Jsoup is yet another HTML parsing library, but it provides a lot of functionalities and boasts much more interesting features when compared to others. Give it a try, and you will see the difference!

Instant jsoup How-to provides simple and detailed instructions on how to use the Jsoup library to manipulate HTML content to suit your needs. You will learn the basic aspects of data crawling, as well as the various concepts of Jsoup so you can make the best use of the library to achieve your goals.

Instant jsoup How-to will help you learn step-by-step using real-world, practical problems. You will begin by learning several basic topics, such as getting input from a URL, a file, or a string, as well as making use of DOM navigation to search for data. You will then move on to some advanced topics like how to use the CSS selector and how to clean dirty HTML data. HTML data is not always safe, and because of that, you will learn how to sanitize the dirty documents to prevent further XSS attacks.

Instant jsoup How-to is a book for every Java developer who wants to learn HTML manipulation quickly and effectively. This book includes the sample source code for you to refer to with a detailed explanation of every feature of the library.

Approach

Filled with practical, step-by-step instructions and clear explanations for the most important and useful tasks. This book will take a how-to approach, focusing on recipes that demonstrate Jsoup.

Who this book is for

If you are working in data scraping, data crawling, or within a similar area using Java, then this book is the one for you. This book acts as a fast-paced and simple guide to enhance your HTML data manipulating skills using one of the most well-known libraries, Jsoup.

Frequently asked questions

How do I cancel my subscription?
Simply head over to the account section in settings and click on “Cancel Subscription” - it’s as simple as that. After you cancel, your membership will stay active for the remainder of the time you’ve paid for. Learn more here.
Can/how do I download books?
At the moment all of our mobile-responsive ePub books are available to download via the app. Most of our PDFs are also available to download and we're working on making the final remaining ones downloadable now. Learn more here.
What is the difference between the pricing plans?
Both plans give you full access to the library and all of Perlego’s features. The only differences are the price and subscription period: With the annual plan you’ll save around 30% compared to 12 months on the monthly plan.
What is Perlego?
We are an online textbook subscription service, where you can get access to an entire online library for less than the price of a single book per month. With over 1 million books across 1000+ topics, we’ve got you covered! Learn more here.
Do you support text-to-speech?
Look out for the read-aloud symbol on your next book to see if you can listen to it. The read-aloud tool reads text aloud for you, highlighting the text as it is being read. You can pause it, speed it up and slow it down. Learn more here.
Is Instant Jsoup How-to an online PDF/ePUB?
Yes, you can access Instant Jsoup How-to by Pete Houston in PDF and/or ePUB format, as well as other popular books in Informatik & Programmierung in HTML. We have over one million books available in our catalogue for you to explore.

Information

Year
2013
ISBN
9781782167990

Instant Jsoup How-to


Instant Jsoup How-to

Copyright © 2013 Packt Publishing
All rights reserved. No part of this book may be reproduced, stored in a retrieval system, or transmitted in any form or by any means, without the prior written permission of the publisher, except in the case of brief quotations embedded in critical articles or reviews.
Every effort has been made in the preparation of this book to ensure the accuracy of the information presented. However, the information contained in this book is sold without warranty, either express or implied. Neither the author, nor Packt Publishing, and its dealers and distributors will be held liable for any damages caused or alleged to be caused directly or indirectly by this book.
Packt Publishing has endeavored to provide trademark information about all of the companies and products mentioned in this book by the appropriate use of capitals. However, Packt Publishing cannot guarantee the accuracy of this information.
First published: June 2013
Production Reference: 1040613
Published by Packt Publishing Ltd.
Livery Place
35 Livery Street
Birmingham B3 2PB, UK.
ISBN 978-1-78216-799-0
www.packtpub.com

Credits

Author
Pete Houston
Reviewers
Pamela Lu
ĂĂ±igo Mediavilla Saiz
Acquisition Editor
Aarthi Kumaraswamy
Commissioning Editor
Poonam Jain
Technical Editor
Dheera Meril Paul
Copy Editor
Brandt D'Mello
Project Coordinator
Suraj Bist
Proofreader
Maria Gould
Production Coordinator
Conidon Miranda
Cover Work
Conidon Miranda
Cover Image
Nitesh Thakur

About the Author

Pete Houston is a software engineer from South Korea with 10 years of experiences in software design and development.
He has undertaken research on medical imaging that helps in diagnosing symptoms of cancer in patients. He has worked with C, C++, COM/DLL, ActiveX Control, and C#.NET 3.0. He also designed and architected the Android mobile platform.
Currently, he is working on the research and implementation of search algorithms for data mining (C, Apache module, Python, and Hadoop).
He has also worked as the technical leader for a backend system to provide information services (Java, Jsoup, PHP, SimpleXML, and Yii-Slim Framework). He also likes to spend his time on sharing technical stuffs on his homepage, http://petehouston.com/.

About the Reviewers

Pamela Lu is a software developer in the Philippines with 9 years of experience. She started working as a programmer in 2004, and since then, most of her projects for work have been on Enterprise web applications. She primarily codes in Java and JavaScript. Outside work, she likes trying out other languages and databases. Currently, she is trying to learn functional programming using Scala.
ĂĂ±igo Mediavilla Saiz is a full stack web developer with three years of experience in web development.
On the backend, he's had the opportunity to work on programming languages for JVM, starting with Java, then moving to Groovy, and finally to Scala. Right now, he is working on the implementation of a highly scalable architecture on top of the Play 2 framework. As a frontend programmer, he has experience in designing single-page apps with Ember, Knockout, and Sammy.
He shares his interests on Scala, Groovy, JavaScript, functional programming, and concurrent and parallel systems and software engineering on his blog, http://imediava.wordpress.com.

www.PacktPub.com

Support files, eBooks, discount offers and more

You might want to visit www.PacktPub.com for support files and downloads related to your book.
Did you know that Packt offers eBook versions of every book published, with PDF and ePub files available? You can upgrade to the eBook version at www.PacktPub.com and as a print book customer, you are entitled to a discount on the eBook copy. Get in touch with us at for more details.
At www.PacktPub.com, you can also read a collection of fre...

Table of contents

Citation styles for Instant Jsoup How-to

APA 6 Citation

Houston, P. (2013). Instant Jsoup How-to (1st ed.). Packt Publishing. Retrieved from https://www.perlego.com/book/389925/instant-jsoup-howto-pdf (Original work published 2013)

Chicago Citation

Houston, Pete. (2013) 2013. Instant Jsoup How-To. 1st ed. Packt Publishing. https://www.perlego.com/book/389925/instant-jsoup-howto-pdf.

Harvard Citation

Houston, P. (2013) Instant Jsoup How-to. 1st edn. Packt Publishing. Available at: https://www.perlego.com/book/389925/instant-jsoup-howto-pdf (Accessed: 14 October 2022).

MLA 7 Citation

Houston, Pete. Instant Jsoup How-To. 1st ed. Packt Publishing, 2013. Web. 14 Oct. 2022.