Budget: 250 USD Deadline: 7 days
Hello, Andrey! Parsing + auto-posting in Telegram is my specialty, and I have already looked at your sources:
— Douban (movie.douban.com/explore) and Ctrip (you.ctrip.com/travels) load content via JavaScript — I will use Playwright (headless browser), not the "light" BeautifulSoup.
— 36Kr (articles/search/profile) provides data in a more structured way, making parsing easier and faster.
What I will do:
• for each material: title, link, date, author/source, brief text, images (where available);
• auto-tags by source + by topic (from the text);
• deduplication through a database (SQLite) — one material will not be posted twice;
• auto-posting to your Telegram channel via a bot on a schedule;
• configuration (sources/tags/frequency), delays, and logging tailored to the specifics of Chinese sites (frequent requests are cut off) — without bypassing protections, only public data.
I suggest starting with an MVP for 2-3 sources (I will check real availability in a few days and provide a working bot), then I will add the rest.
Questions: should I publish immediately upon finding or through moderation? Is a translation/summary of the Chinese text needed before posting? Tags only by source or also by topics within the text?
$250 for a working MVP (2-3 sources + bot + dedup + deploy), further sources — we will negotiate. Ready to start immediately.
Budget: 2000 USD Deadline: 7 days
Hello, I worked on a news parser for a media aggregator — collecting articles from 12 sources, ~500 publications/day, auto-posting to Telegram with tags and deduplication.
Regarding your project: some of the mentioned sites (for example, Douban and Ctrip) use dynamic loading via JavaScript — are you open to using Playwright instead of lightweight BeautifulSoup if the site requires it?
I suggest we get in touch; I will provide you with free technical consultation and we can create a development plan + I will tell you about my team!
Budget: 15000 USD Deadline: 7 days
Good day! We have experience in developing parsers in Python with bypassing protection and integration with the Telegram API. We implement this through the Playwright library for dynamic content and asynchronous message sending to the channel. We will set up stable script operation on the server, taking into account the specifics of Chinese resources.
Maksym Potashov
Winning proposal- Projects 6
- Rating 3.9
- Rating 788
Budget: 150 USD Deadline: 7 days
I would start by checking each source separately: Douban, Ctrip, and 36Kr may deliver pages differently, so I will first determine where regular parsing is sufficient and where dynamic loading processing is needed. For the MVP, I would use Python, Playwright/BeautifulSoup, SQLite, Telegram Bot API, and Docker to ensure everything can run smoothly on the server. Duplicates will be stored in the database, and the check frequency and tags will be in the config.
I have worked with Chinese websites, where often the headache is not in the parsing itself, but in the fact that some data is loaded separately or the site may throttle frequent requests. Therefore, I will immediately incorporate delays, logging, and a brief description of the limitations for each site, without bypassing protections.
An MVP for 2-3 sources can be assembled in a few days, and the full version for all sources after checking availability. I will provide a more accurate budget after a quick technical review of the sites.
Please advise, should the publication on Telegram go out immediately after finding the material or go through moderation? Are tags needed only for the source or also for topics within the text? And should the Chinese text be translated/shortened before publication?
Proposals are currently absent
Current freelance projects in the category Data Parsing
I'm looking for a specialist for a one-time task of mass messaging on Telegram. I need someone who can take the process on themselves and execute it "turnkey". What needs to be done: - Prepare the account base: create/purchase and set up a network of working accounts for messaging. - Gather the audience (Parsing): parse active users from thematic chats. - Send the messages: send messages to the gathered base in personal messages (DMs). Requirements for execution: - You have the necessary software for parsing and messaging. - You understand the current limits of Telegram, know how to work with proxies, text randomization, and know how to minimize account blocking. In your response, please write: - What software do you use for work? - What is the cost of your services (indicate the price for the entire volume of work or per 1000 successfully delivered messages)? - How much time will you need for preparation and launch?
A parser needs to be created to collect competitor prices, with a catalog of over 40,000 products. Access to the prices is only available under authorization.
Hello I previously posted a project (description below) and received the following result https://docs.google.com/spreadsheets/d/1-abSBtjRZeD_ZGriquW0YUZdbmC4eXcr/edit?usp=sharing&ouid=113724943543706943348&rtpof=true&sd=true as it turned out, unfortunately, later - it is very incomplete... for example, in the table, there is a filter for the city of Neperville - where only 4 dealers are listed. But I manually found 25 such dealers on the maps, and these are not duplicates but unique. Some cities are completely missing. So I am interested in the city of Chicago and the surrounding towns (suburbs), I will provide a list if needed. ---- List of categories from Google Maps from which all companies need to be extracted Car Dealer Used Car Dealer New Car Dealer I also provide a list of categories used by branded dealerships: Toyota Dealer Ford Dealer BMW Dealer Honda Dealer I will provide a list of car brands separately. Generally, within the city, there are several dealerships for each car brand The result in the form of a table. Deduplication through Place ID Price for the work. This is a one-time job. I do not require any script that I will work with further. You can do parsing through your developments or third-party ready services.
It is necessary to set up the parser remotely (via Anydesk or TeamViewer), something is broken and not working correctly. Mainly replacing the starting links and configuring the parser. Website https://dveri-pol.com/
A modern online toy store needs to be created on WordPress (WooCommerce). Main requirements: - Responsive design for PCs, tablets, and smartphones. - Product catalog with categories, filters, and search. - Product pages with descriptions, photos, specifications, and prices. - Cart and checkout. - Integration of popular payment and delivery methods. - Customer personal account. - User-friendly control panel for adding and editing products. - Basic SEO optimization and high website loading speed. - Security, backup, and SSL settings. It is preferable to use a modern, lightweight theme or create a custom design. After the work is completed, a brief management guide for the website must be provided.