c#网页抓取京东商城首页案列-阿里云开发者社区

c#网页抓取京东商城首页案列

2024-05-08 156

版权

本文内容由阿里云实名注册用户自发贡献，版权归原作者所有，阿里云开发者社区不拥有其著作权，亦不承担相应法律责任。具体规则请查看《阿里云开发者社区用户服务协议》和《阿里云开发者社区知识产权保护指引》。如果您发现本社区中有涉嫌抄袭的内容，填写侵权投诉表单进行举报，一经查实，本社区将立刻删除涉嫌侵权内容。

简介： c#网页抓取京东商城首页案列

在C#语言中，我们可以使用一些工具类库来进行网页抓取，例如.NET Framework提供的System.Net命名空间中的HttpWebRequest和HttpWebResponse类，以及第三方类库HtmlAgilityPack。

下面是一个简单的C#程序，使用HttpWebRequest和HttpWebResponse类来抓取网页：

using System;
using System.IO;
using System.Net;
 
public class WebRequestExample
{
    public static void Main()
    {
        // 创建一个WebRequest对象
        WebRequest request = WebRequest.Create("http://www.example.com");
 
        // 发送请求并获取响应
        WebResponse response = request.GetResponse();
 
        // 读取响应内容
        Stream stream = response.GetResponseStream();
        StreamReader reader = new StreamReader(stream);
        string content = reader.ReadToEnd();
        
        // 关闭响应和流
        response.Close();
        reader.Close();
        stream.Close();
 
        // 输出网页内容
        Console.WriteLine(content);
    }
}

上述代码会获取http://www.example.com的网页内容并输出到控制台中。

如果需要解析网页内容，可以使用HtmlAgilityPack类库。下面是一个使用HtmlAgilityPack来抓取京东商城首页商品列表的示例：

using System;
using System.Collections.Generic;
using System.Linq;
using System.Net.Http;
using HtmlAgilityPack;
 
public class Program
{
    public static void Main()
    {
        // 发送HTTP请求获取京东商城首页内容
        using (var client = new HttpClient())
        {
            var response = client.GetAsync("https://www.jd.com/").Result;
            var responseContent = response.Content;
 
            // 使用HtmlAgilityPack解析HTML内容
            var htmlDocument = new HtmlDocument();
            htmlDocument.LoadHtml(responseContent.ReadAsStringAsync().Result);
 
            // 获取商品列表中的商品名称和价格
            var productList = htmlDocument.DocumentNode.Descendants("ul")
                .Where(x => x.GetAttributeValue("class", "") == "clearfix")
                .FirstOrDefault();
            if (productList != null)
            {
                var items = productList.Descendants("li").ToList();
                foreach (var item in items)
                {
                    var name = item.Descendants("div")
                        .Where(x => x.GetAttributeValue("class", "") == "p-name")
                        .FirstOrDefault()?.InnerText;
                    var price = item.Descendants("div")
                        .Where(x => x.GetAttributeValue("class", "") == "p-price")
                        .FirstOrDefault()?.InnerText;
                    Console.WriteLine($"商品名称：{name}，价格：{price}");
                }
            }
        }
    }
}

上述代码使用HttpClient类发送HTTP请求获取京东商城首页内容，然后使用HtmlAgilityPack解析HTML内容，并提取出商品列表中的商品名称和价格。最后输出到控制台中。

需要注意的是，网页抓取涉及到网站的隐私和版权等问题，需要遵循相关法律法规，不得擅自抓取或使用他人内容。

文章标签：

c#网页抓取京东商城首页案列

热门文章

最新文章

相关电子书

探索云世界

热门

云计算

大数据

云原生

人工智能

数据库

开发与运维

活动广场

任务中心

训练营

直播

乘风者计划

下载

镜像站

技术资料

c#网页抓取京东商城首页案列

热门文章

最新文章

相关电子书