Zig
Zig
Fetch robots.txt for a Site
See more Spider Examples
The Chilkat Spider library is robots.txt compliant. It automatically fetches a site's robots.txt file and adheres to it. It will not download pages denied by robots.txt. Pages excluded by robots.txt will not appear in the Spider's "unspidered" list. This example shows how to explicitly download and review the robots.txt for a given site.Chilkat Zig Downloads
const std = @import("std");
const chilkat = @import("chilkat");
pub fn main(init: std.process.Init) !void {
const alloc = init.arena.allocator();
const spider = try chilkat.Spider.init();
defer spider.deinit();
spider.initialize("www.chilkatsoft.com");
const robots_text = try spider.fetchRobotsText(alloc);
std.debug.print("{s}\n", .{robots_text});
}