Home Backend Development PHP Tutorial PHP programmers use crawler technology to reveal the real data behind rising rents

PHP programmers use crawler technology to reveal the real data behind rising rents

Aug 25, 2018 am 11:46 AM

In the near future, I believe that everyone on Weibo or in the circle of friends has been flooded by the skyrocketing rent, the vice president of I Love My Family published a resignation letter in the circle of friends announcing his resignation, and the Internet exposure of Lianjia Ziroom’s jacking up of housing prices, etc. Pass. When rents rise, young people are most affected. Most young people either have just graduated and have no savings, or they have worked for several years but continue to struggle to rent a house due to high housing prices. Now even renting a house has become a big problem. So as a PHP programmer, let me introduce this incident to you how to use PHP to write a crawler to obtain real rental data.

PHP programmers use crawler technology to reveal the real data behind rising rents

Regarding the rental market in Beijing, if you want to rent a house, there are three main ways: 1. Find a housing agency. The company with the highest market share currently is called Lianjia; 2. The one with the highest market share for long-term rental apartments is called Ziru; 3. The one with the highest market share for finding apartments on the Internet is Anjuke. In April this year, there was a new company that suddenly emerged and quickly jumped into the top five, called Beike House Search. The sum of these three methods almost determines the price of renting a house for you and me. What is even more surprising is that the above-mentioned companies , except for Anjuke, the actual controller of Lianjia, Ziroom, and Beikezhuofang is the same person. This is Zuo Hui, the boss of Lianjia Group who has frequently appeared in the news these days.

PHP programmers use crawler technology to reveal the real data behind rising rents

#For those who are preparing to work hard in Beijing, the skyrocketing rent is quite annoying. Some netizens used programmers’ methods to uncover the reasons behind the increase in rent. So what is the programmer's way?

The program idea is: Write a crawler with phpUse it to crawl the data of Lianjia. First, go to the console to see the loading information, find the relevant data API, send an https request according to the required parameters in the request header, and after the analysis is completed, use xpath or regular expression tools to match the content you want, and then insert it into the database, that is Fetching can be completed. Finally PHP implemented crawler crawled all the houses for rent on Lianjia.com.

PHP programmers use crawler technology to reveal the real data behind rising rents

Then continue to use the same crawler method to crawl long-term rental apartment platforms such as Ziru, Danke, and Mushroom Apartments. The final word cloud diagram of the data is as follows

PHP programmers use crawler technology to reveal the real data behind rising rents

According to the data summary, in several major directions of Beijing’s rental industry, Zuobobo’s industry either occupies a leading position or is growing rapidly. It’s no wonder that there was an article a few days ago Heavy news said that Hu Jinghui, the former vice president of I Love My Home, resigned due to some pressure and criticized long-term rental apartments such as Ziru and Danke for competing for housing at prices 20%-40% higher than the market price, regardless of costs. to expand.

PHP programmers use crawler technology to reveal the real data behind rising rents

It is understandable for businessmen to pursue profit, and it is also a normal business goal to pursue a larger market share. However, when a certain enterprise is too strong, a monopoly or oligopoly will be formed. If they have a monopoly, they can use their resources and capital advantages to hoard, influence and even manipulate the direction of the industry. Such a monopoly seems to be taking shape in Beijing's rental industry. The main purpose here is to tell you that PHP crawler can obtain network resources such as web pages, pictures, scripts, file data, etc. from the Internet.


The above is the detailed content of PHP programmers use crawler technology to reveal the real data behind rising rents. For more information, please follow other related articles on the PHP Chinese website!

Statement of this Website
The content of this article is voluntarily contributed by netizens, and the copyright belongs to the original author. This site does not assume corresponding legal responsibility. If you find any content suspected of plagiarism or infringement, please contact admin@php.cn

Hot AI Tools

Undresser.AI Undress

Undresser.AI Undress

AI-powered app for creating realistic nude photos

AI Clothes Remover

AI Clothes Remover

Online AI tool for removing clothes from photos.

Undress AI Tool

Undress AI Tool

Undress images for free

Clothoff.io

Clothoff.io

AI clothes remover

Video Face Swap

Video Face Swap

Swap faces in any video effortlessly with our completely free AI face swap tool!

Hot Tools

Notepad++7.3.1

Notepad++7.3.1

Easy-to-use and free code editor

SublimeText3 Chinese version

SublimeText3 Chinese version

Chinese version, very easy to use

Zend Studio 13.0.1

Zend Studio 13.0.1

Powerful PHP integrated development environment

Dreamweaver CS6

Dreamweaver CS6

Visual web development tools

SublimeText3 Mac version

SublimeText3 Mac version

God-level code editing software (SublimeText3)

Explain JSON Web Tokens (JWT) and their use case in PHP APIs. Explain JSON Web Tokens (JWT) and their use case in PHP APIs. Apr 05, 2025 am 12:04 AM

JWT is an open standard based on JSON, used to securely transmit information between parties, mainly for identity authentication and information exchange. 1. JWT consists of three parts: Header, Payload and Signature. 2. The working principle of JWT includes three steps: generating JWT, verifying JWT and parsing Payload. 3. When using JWT for authentication in PHP, JWT can be generated and verified, and user role and permission information can be included in advanced usage. 4. Common errors include signature verification failure, token expiration, and payload oversized. Debugging skills include using debugging tools and logging. 5. Performance optimization and best practices include using appropriate signature algorithms, setting validity periods reasonably,

What are Enumerations (Enums) in PHP 8.1? What are Enumerations (Enums) in PHP 8.1? Apr 03, 2025 am 12:05 AM

The enumeration function in PHP8.1 enhances the clarity and type safety of the code by defining named constants. 1) Enumerations can be integers, strings or objects, improving code readability and type safety. 2) Enumeration is based on class and supports object-oriented features such as traversal and reflection. 3) Enumeration can be used for comparison and assignment to ensure type safety. 4) Enumeration supports adding methods to implement complex logic. 5) Strict type checking and error handling can avoid common errors. 6) Enumeration reduces magic value and improves maintainability, but pay attention to performance optimization.

How does session hijacking work and how can you mitigate it in PHP? How does session hijacking work and how can you mitigate it in PHP? Apr 06, 2025 am 12:02 AM

Session hijacking can be achieved through the following steps: 1. Obtain the session ID, 2. Use the session ID, 3. Keep the session active. The methods to prevent session hijacking in PHP include: 1. Use the session_regenerate_id() function to regenerate the session ID, 2. Store session data through the database, 3. Ensure that all session data is transmitted through HTTPS.

Describe the SOLID principles and how they apply to PHP development. Describe the SOLID principles and how they apply to PHP development. Apr 03, 2025 am 12:04 AM

The application of SOLID principle in PHP development includes: 1. Single responsibility principle (SRP): Each class is responsible for only one function. 2. Open and close principle (OCP): Changes are achieved through extension rather than modification. 3. Lisch's Substitution Principle (LSP): Subclasses can replace base classes without affecting program accuracy. 4. Interface isolation principle (ISP): Use fine-grained interfaces to avoid dependencies and unused methods. 5. Dependency inversion principle (DIP): High and low-level modules rely on abstraction and are implemented through dependency injection.

Explain late static binding in PHP (static::). Explain late static binding in PHP (static::). Apr 03, 2025 am 12:04 AM

Static binding (static::) implements late static binding (LSB) in PHP, allowing calling classes to be referenced in static contexts rather than defining classes. 1) The parsing process is performed at runtime, 2) Look up the call class in the inheritance relationship, 3) It may bring performance overhead.

What is REST API design principles? What is REST API design principles? Apr 04, 2025 am 12:01 AM

RESTAPI design principles include resource definition, URI design, HTTP method usage, status code usage, version control, and HATEOAS. 1. Resources should be represented by nouns and maintained at a hierarchy. 2. HTTP methods should conform to their semantics, such as GET is used to obtain resources. 3. The status code should be used correctly, such as 404 means that the resource does not exist. 4. Version control can be implemented through URI or header. 5. HATEOAS boots client operations through links in response.

How do you handle exceptions effectively in PHP (try, catch, finally, throw)? How do you handle exceptions effectively in PHP (try, catch, finally, throw)? Apr 05, 2025 am 12:03 AM

In PHP, exception handling is achieved through the try, catch, finally, and throw keywords. 1) The try block surrounds the code that may throw exceptions; 2) The catch block handles exceptions; 3) Finally block ensures that the code is always executed; 4) throw is used to manually throw exceptions. These mechanisms help improve the robustness and maintainability of your code.

What are anonymous classes in PHP and when might you use them? What are anonymous classes in PHP and when might you use them? Apr 04, 2025 am 12:02 AM

The main function of anonymous classes in PHP is to create one-time objects. 1. Anonymous classes allow classes without names to be directly defined in the code, which is suitable for temporary requirements. 2. They can inherit classes or implement interfaces to increase flexibility. 3. Pay attention to performance and code readability when using it, and avoid repeatedly defining the same anonymous classes.

See all articles