


PHP programmers use crawler technology to reveal the real data behind rising rents
In the near future, I believe that everyone on Weibo or in the circle of friends has been flooded by the skyrocketing rent, the vice president of I Love My Family published a resignation letter in the circle of friends announcing his resignation, and the Internet exposure of Lianjia Ziroom’s jacking up of housing prices, etc. Pass. When rents rise, young people are most affected. Most young people either have just graduated and have no savings, or they have worked for several years but continue to struggle to rent a house due to high housing prices. Now even renting a house has become a big problem. So as a PHP programmer, let me introduce this incident to you how to use PHP to write a crawler to obtain real rental data.
Regarding the rental market in Beijing, if you want to rent a house, there are three main ways: 1. Find a housing agency. The company with the highest market share currently is called Lianjia; 2. The one with the highest market share for long-term rental apartments is called Ziru; 3. The one with the highest market share for finding apartments on the Internet is Anjuke. In April this year, there was a new company that suddenly emerged and quickly jumped into the top five, called Beike House Search. The sum of these three methods almost determines the price of renting a house for you and me. What is even more surprising is that the above-mentioned companies , except for Anjuke, the actual controller of Lianjia, Ziroom, and Beikezhuofang is the same person. This is Zuo Hui, the boss of Lianjia Group who has frequently appeared in the news these days.
#For those who are preparing to work hard in Beijing, the skyrocketing rent is quite annoying. Some netizens used programmers’ methods to uncover the reasons behind the increase in rent. So what is the programmer's way?
The program idea is: Write a crawler with phpUse it to crawl the data of Lianjia. First, go to the console to see the loading information, find the relevant data API, send an https request according to the required parameters in the request header, and after the analysis is completed, use xpath or regular expression tools to match the content you want, and then insert it into the database, that is Fetching can be completed. Finally PHP implemented crawler crawled all the houses for rent on Lianjia.com.
Then continue to use the same crawler method to crawl long-term rental apartment platforms such as Ziru, Danke, and Mushroom Apartments. The final word cloud diagram of the data is as follows
According to the data summary, in several major directions of Beijing’s rental industry, Zuobobo’s industry either occupies a leading position or is growing rapidly. It’s no wonder that there was an article a few days ago Heavy news said that Hu Jinghui, the former vice president of I Love My Home, resigned due to some pressure and criticized long-term rental apartments such as Ziru and Danke for competing for housing at prices 20%-40% higher than the market price, regardless of costs. to expand.
It is understandable for businessmen to pursue profit, and it is also a normal business goal to pursue a larger market share. However, when a certain enterprise is too strong, a monopoly or oligopoly will be formed. If they have a monopoly, they can use their resources and capital advantages to hoard, influence and even manipulate the direction of the industry. Such a monopoly seems to be taking shape in Beijing's rental industry. The main purpose here is to tell you that PHP crawler can obtain network resources such as web pages, pictures, scripts, file data, etc. from the Internet.
The above is the detailed content of PHP programmers use crawler technology to reveal the real data behind rising rents. For more information, please follow other related articles on the PHP Chinese website!

Hot AI Tools

Undresser.AI Undress
AI-powered app for creating realistic nude photos

AI Clothes Remover
Online AI tool for removing clothes from photos.

Undress AI Tool
Undress images for free

Clothoff.io
AI clothes remover

Video Face Swap
Swap faces in any video effortlessly with our completely free AI face swap tool!

Hot Article

Hot Tools

Notepad++7.3.1
Easy-to-use and free code editor

SublimeText3 Chinese version
Chinese version, very easy to use

Zend Studio 13.0.1
Powerful PHP integrated development environment

Dreamweaver CS6
Visual web development tools

SublimeText3 Mac version
God-level code editing software (SublimeText3)

Hot Topics

JWT is an open standard based on JSON, used to securely transmit information between parties, mainly for identity authentication and information exchange. 1. JWT consists of three parts: Header, Payload and Signature. 2. The working principle of JWT includes three steps: generating JWT, verifying JWT and parsing Payload. 3. When using JWT for authentication in PHP, JWT can be generated and verified, and user role and permission information can be included in advanced usage. 4. Common errors include signature verification failure, token expiration, and payload oversized. Debugging skills include using debugging tools and logging. 5. Performance optimization and best practices include using appropriate signature algorithms, setting validity periods reasonably,

Session hijacking can be achieved through the following steps: 1. Obtain the session ID, 2. Use the session ID, 3. Keep the session active. The methods to prevent session hijacking in PHP include: 1. Use the session_regenerate_id() function to regenerate the session ID, 2. Store session data through the database, 3. Ensure that all session data is transmitted through HTTPS.

The enumeration function in PHP8.1 enhances the clarity and type safety of the code by defining named constants. 1) Enumerations can be integers, strings or objects, improving code readability and type safety. 2) Enumeration is based on class and supports object-oriented features such as traversal and reflection. 3) Enumeration can be used for comparison and assignment to ensure type safety. 4) Enumeration supports adding methods to implement complex logic. 5) Strict type checking and error handling can avoid common errors. 6) Enumeration reduces magic value and improves maintainability, but pay attention to performance optimization.

The application of SOLID principle in PHP development includes: 1. Single responsibility principle (SRP): Each class is responsible for only one function. 2. Open and close principle (OCP): Changes are achieved through extension rather than modification. 3. Lisch's Substitution Principle (LSP): Subclasses can replace base classes without affecting program accuracy. 4. Interface isolation principle (ISP): Use fine-grained interfaces to avoid dependencies and unused methods. 5. Dependency inversion principle (DIP): High and low-level modules rely on abstraction and are implemented through dependency injection.

Static binding (static::) implements late static binding (LSB) in PHP, allowing calling classes to be referenced in static contexts rather than defining classes. 1) The parsing process is performed at runtime, 2) Look up the call class in the inheritance relationship, 3) It may bring performance overhead.

RESTAPI design principles include resource definition, URI design, HTTP method usage, status code usage, version control, and HATEOAS. 1. Resources should be represented by nouns and maintained at a hierarchy. 2. HTTP methods should conform to their semantics, such as GET is used to obtain resources. 3. The status code should be used correctly, such as 404 means that the resource does not exist. 4. Version control can be implemented through URI or header. 5. HATEOAS boots client operations through links in response.

In PHP, exception handling is achieved through the try, catch, finally, and throw keywords. 1) The try block surrounds the code that may throw exceptions; 2) The catch block handles exceptions; 3) Finally block ensures that the code is always executed; 4) throw is used to manually throw exceptions. These mechanisms help improve the robustness and maintainability of your code.

The main function of anonymous classes in PHP is to create one-time objects. 1. Anonymous classes allow classes without names to be directly defined in the code, which is suitable for temporary requirements. 2. They can inherit classes or implement interfaces to increase flexibility. 3. Pay attention to performance and code readability when using it, and avoid repeatedly defining the same anonymous classes.
