CWE-79: Vô hiệu hóa không đúng dữ liệu đầu vào khi tạo trang web (Cross-site Scripting)

CWE-79 là gì?

Sản phẩm không vô hiệu hóa, hoặc vô hiệu hóa không đúng, dữ liệu do người dùng kiểm soát trước khi đưa dữ liệu đó vào nội dung trang web phân phối cho người dùng khác. Vì vậy, trình duyệt có thể diễn giải dữ liệu nguy hiểm như nội dung chủ động.

Thống kê dữ liệu

THỨ HẠNG OWASP TOP 10:20255 — A05:2025 — Injection
TỔNG SỐ CVE LIÊN QUAN (365 NGÀY)3.518
MỨC TRỪU TƯỢNGCơ bản
KHẢ NĂNG KHAI THÁCCao

Số lượng lỗ hổng nằm trong CWE-79

3.518 lỗ hổngTăng 106,5% so với cùng kỳ

Số lượng lỗ hổng trong CISA KEV của CWE-79

3 lỗ hổngTăng 200% so với cùng kỳ

Định nghĩa chính thức

TheoMitre CWE

Chi tiết kỹ thuật

Cross-site scripting, thường viết tắt là XSS, là một nhóm điểm yếu có nhiều đường truyền và mô hình tấn công khác nhau, nhưng đều bắt nguồn từ cùng một sai sót: dữ liệu nguy hiểm không được vô hiệu hóa đúng cách giữa kẻ tấn công và nạn nhân. Điểm yếu có thể xảy ra khi dữ liệu được phản hồi trực tiếp từ yêu cầu HTTP, được lưu trong kho dữ liệu tin cậy rồi hiển thị sau đó, hoặc được ứng dụng phía máy khách chèn vào thông qua DOM. Cơ chế same-origin của trình duyệt nhằm tách tài nguyên phía máy khách của một trang khỏi tài nguyên của origin khác, nhưng mã lệnh được chèn có thể chạy trong ngữ cảnh của trang bị ảnh hưởng và truy cập các tài nguyên mà origin đó được phép sử dụng.

Đặc điểm

Đây là một điểm yếu cơ sở, đơn giản, ổn định, thường phát sinh trong giai đoạn triển khai và không phụ thuộc vào ngôn ngữ lập trình cụ thể. Các dạng phổ biến gồm XSS phản chiếu hoặc không liên tục, XSS lưu trữ hoặc liên tục, và XSS dựa trên DOM. Dữ liệu không tin cậy có thể đi vào qua tham số, cookie, dữ liệu mạng, biến môi trường, kết quả tra cứu DNS ngược, kết quả truy vấn, tiêu đề yêu cầu, thành phần URL, email, tệp, tên tệp, cơ sở dữ liệu, hệ thống bên ngoài hoặc các lời gọi API gián tiếp. XSS thường gặp trong công nghệ dựa trên web; bản ghi cũng xác định máy chủ web và công nghệ AI/ML là các nền tảng áp dụng.

Hậu quả thường gặp

XSS có thể ảnh hưởng đến tính bảo mật, tính toàn vẹn, tính sẵn sàng và kiểm soát truy cập. Mã lệnh có thể đọc dữ liệu ứng dụng như thông tin phiên trong cookie, thực hiện hành động thay mặt nạn nhân, vượt qua cơ chế bảo vệ, làm lộ thông tin mật, chuyển hướng người dùng, thay đổi cách hiển thị trang, làm lộ tệp của người dùng cuối, cài đặt chương trình Trojan hoặc, khi kết hợp với điểm yếu khác, chạy mã độc. Mức độ ảnh hưởng có thể từ gây phiền toái đến chiếm quyền tài khoản hoàn toàn; rủi ro đặc biệt cao khi nạn nhân có quyền quản trị. Trong một số trường hợp, mã tùy ý còn có thể chạy trên máy tính của nạn nhân.

Dữ liệu MITRE CWE chính thức
Tác độngPhạm viDiễn giải
Vượt qua cơ chế bảo vệ, Đọc dữ liệu ứng dụngKiểm soát truy cập, Tính bí mậtThe most common attack performed with cross-site scripting involves the disclosure of private information stored in user cookies, such as session information. Typically, a malicious user will craft a client-side script, which -- when parsed by a web browser -- performs some activity on behalf of the victim to an attacker-controlled system (such as sending all site cookies to a given E-mail address). This could be especially dangerous to the site if the victim has administrator privileges to manage that site. This script will be loaded and run by each user visiting the web site. Since the site requesting to run the script has access to the cookies in question, the malicious script does also.
Thực thi mã hoặc lệnh trái phépTính toàn vẹn, Tính bí mật, Tính sẵn sàngIn some circumstances it may be possible to run arbitrary code on a victim's computer when cross-site scripting is combined with other flaws, for example, "drive-by hacking."
Thực thi mã hoặc lệnh trái phép, Vượt qua cơ chế bảo vệ, Đọc dữ liệu ứng dụngTính bí mật, Tính toàn vẹn, Tính sẵn sàng, Kiểm soát truy cậpThe consequence of an XSS attack is the same regardless of whether it is stored or reflected. The difference is in how the payload arrives at the server. XSS can cause a variety of problems for the end user that range in severity from an annoyance to complete account compromise. Some cross-site scripting vulnerabilities can be exploited to manipulate or steal cookies, create requests that can be mistaken for those of a valid user, compromise confidential information, or execute malicious code on the end user systems for a variety of nefarious purposes. Other damaging attacks include the disclosure of end user files, installation of Trojan horse programs, redirecting the user to some other page or site, running "Active X" controls (under Microsoft Internet Explorer) from sites that a user perceives as trustworthy, and modifying presentation of content.

Biện pháp giảm thiểu rủi ro

Xử lý đầu ra theo đúng ngữ cảnh. Ưu tiên thư viện hoặc framework đã được thẩm định có sẵn cơ chế mã hóa đầu ra an toàn. Xác định kiểu mã hóa cần thiết cho từng ngữ cảnh đầu ra, gồm thân HTML, thuộc tính phần tử, URI, phần JavaScript và CSS hoặc thuộc tính style; mã hóa thực thể HTML chỉ phù hợp với thân HTML. Chỉ định kiểu mã hóa mà thành phần downstream sẽ xử lý nhất quán, chẳng hạn ISO-8859-1, UTF-7 hoặc UTF-8, thay vì để trình duyệt hay thành phần khác tự phỏng đoán.

Giới hạn và kiểm tra dữ liệu đầu vào. Coi mọi dữ liệu đầu vào là độc hại và dùng allowlist nghiêm ngặt dựa trên kiểu, độ dài, giá trị, cú pháp, tính nhất quán và quy tắc nghiệp vụ dự kiến. Kiểm tra và làm sạch toàn bộ dữ liệu trong yêu cầu, gồm trường ẩn, cookie, tiêu đề, URL và các trường khác dù ứng dụng không dự kiến hiển thị lại; không chỉ dựa vào denylist. Xác định mọi nguồn đầu vào có thể có và dùng cơ chế có cấu trúc để tách dữ liệu khỏi mã. Khi tập tên tệp hoặc URL hợp lệ bị giới hạn, ánh xạ các giá trị đầu vào cố định như ID số tới đối tượng tương ứng và từ chối giá trị khác. Lặp lại các kiểm tra bảo mật phía máy khách ở phía máy chủ. Với Struts, đặt thuộc tính filter của bean thành true cho dữ liệu form bean; với PHP, không dùng register_globals hoặc cơ chế mô phỏng không an toàn.

Tăng cường phòng thủ. Đặt cookie phiên thành HttpOnly khi trình duyệt hỗ trợ, nhưng cần hiểu rằng biện pháp này không ngăn được mọi XSS và không được mọi trình duyệt hỗ trợ. Tường lửa ứng dụng có thể là biện pháp tạm thời hoặc bổ sung khi chưa thể sửa mã ngay, nhưng có thể bỏ sót vector đầu vào, bị vượt qua hoặc từ chối hay sửa đổi yêu cầu hợp lệ.

Dữ liệu MITRE CWE chính thức
  1. Thư viện hoặc framework · Kiến trúc và thiết kếUse a vetted library or framework that does not allow this weakness to occur or provides constructs that make this weakness easier to avoid [REF-1482]. Examples of libraries and frameworks that make it easier to generate properly encoded output include Microsoft's Anti-XSS library, the OWASP ESAPI Encoding module, and Apache Wicket.
  2. Hiện thực hóa, Kiến trúc và thiết kếUnderstand the context in which your data will be used and the encoding that will be expected. This is especially important when transmitting data between different components, or when generating outputs that can contain multiple encodings at the same time, such as web pages or multi-part mail messages. Study all expected communication protocols and data representations to determine the required encoding strategies. For any data that will be output to another web page, especially any data that was received from external inputs, use the appropriate encoding on all non-alphanumeric characters. Parts of the same output document may require different encodings, which will vary depending on whether the output is in the: - HTML body - Element attributes (such as src="XYZ") - URIs - JavaScript sections - Cascading Style Sheets and style property etc. Note that HTML Entity Encoding is only appropriate for the HTML body. Consult the XSS Prevention Cheat Sheet [REF-724] for more details on the types of encoding and escaping that are needed.
  3. Giảm bề mặt tấn công · Kiến trúc và thiết kế, Hiện thực hóa · Hiệu quả: Hạn chếUnderstand all the potential areas where untrusted inputs can enter your software: parameters or arguments, cookies, anything read from the network, environment variables, reverse DNS lookups, query results, request headers, URL components, e-mail, files, filenames, databases, and any external systems that provide data to the application. Remember that such inputs may be obtained indirectly through API calls.This technique has limited effectiveness, but can be helpful when it is possible to store client state and sensitive information on the server side instead of in cookies, headers, hidden form fields, etc.
  4. Kiến trúc và thiết kếFor any security checks that are performed on the client side, ensure that these checks are duplicated on the server side, in order to avoid CWE-602. Attackers can bypass the client-side checks by modifying values after the checks have been performed, or by changing the client to remove the client-side checks entirely. Then, these modified values would be submitted to the server.
  5. Tham số hóa · Kiến trúc và thiết kếIf available, use structured mechanisms that automatically enforce the separation between data and code. These mechanisms may be able to provide the relevant quoting, encoding, and validation automatically, instead of relying on the developer to provide this capability at every point where output is generated.
  6. Mã hóa ký tự đầu ra · Hiện thực hóaUse and specify an output encoding that can be handled by the downstream component that is reading the output. Common encodings include ISO-8859-1, UTF-7, and UTF-8. When an encoding is not specified, a downstream component may choose a different encoding, either by assuming a default encoding or automatically inferring which encoding is being used, which can be erroneous. When the encodings are inconsistent, the downstream component might treat some character or byte sequences as special, even if they are not special in the original encoding. Attackers might then be able to exploit this discrepancy and conduct injection attacks; they even might be able to bypass protection mechanisms that assume the original encoding is also being used by the downstream component. The problem of inconsistent output encodings often arises in web pages. If an encoding is not specified in an HTTP header, web browsers often guess about which encoding is being used. This can open up the browser to subtle XSS attacks.
  7. Hiện thực hóaWith Struts, write all data from form beans with the bean's filter attribute set to true.
  8. Giảm bề mặt tấn công · Hiện thực hóa · Hiệu quả: Phòng thủ nhiều lớpTo help mitigate XSS attacks against the user's session cookie, set the session cookie to be HttpOnly. In browsers that support the HttpOnly feature (such as more recent versions of Internet Explorer and Firefox), this attribute can prevent the user's session cookie from being accessible to malicious client-side scripts that use document.cookie. This is not a complete solution, since HttpOnly is not supported by all browsers. More importantly, XmlHttpRequest and other powerful browser technologies provide read access to HTTP headers, including the Set-Cookie header in which the HttpOnly flag is set.
  9. Kiểm tra dữ liệu đầu vào · Hiện thực hóaAssume all input is malicious. Use an "accept known good" input validation strategy, i.e., use a list of acceptable inputs that strictly conform to specifications. Reject any input that does not strictly conform to specifications, or transform it into something that does. When performing input validation, consider all potentially relevant properties, including length, type of input, the full range of acceptable values, missing or extra inputs, syntax, consistency across related fields, and conformance to business rules. As an example of business rule logic, "boat" may be syntactically valid because it only contains alphanumeric characters, but it is not valid if the input is only expected to contain colors such as "red" or "blue." Do not rely exclusively on looking for malicious or malformed inputs. This is likely to miss at least one undesirable input, especially if the code's environment changes. This can give attackers enough room to bypass the intended validation. However, denylists can be useful for detecting potential attacks or determining which inputs are so malformed that they should be rejected outright. When dynamically constructing web pages, use stringent allowlists that limit the character set based on the expected value of the parameter in the request. All input should be validated and cleansed, not just parameters that the user is supposed to specify, but all data in the request, including hidden fields, cookies, headers, the URL itself, and so forth. A common mistake that leads to continuing XSS vulnerabilities is to validate only fields that are expected to be redisplayed by the site. It is common to see data from the request that is reflected by the application server or the application that the development team did not anticipate. Also, a field that is not currently reflected may be used by a future developer. Therefore, validating ALL parts of the HTTP request is recommended. Note that proper output encoding, escaping, and quoting is the most effective solution for preventing XSS, although input validation may provide some defense-in-depth. This is because it effectively limits what will appear in output. Input validation will not always prevent XSS, especially if you are required to support free-form text fields that could contain arbitrary characters. For example, in a chat application, the heart emoticon ("<3") would likely pass the validation step, since it is commonly used. However, it cannot be directly inserted into the web page because it contains the "<" character, which would need to be escaped or otherwise handled. In this case, stripping the "<" might reduce the risk of XSS, but it would produce incorrect behavior because the emoticon would not be recorded. This might seem to be a minor inconvenience, but it would be more important in a mathematical forum that wants to represent inequalities. Even if you make a mistake in your validation (such as forgetting one out of 100 input fields), appropriate encoding is still likely to protect you from injection-based attacks. As long as it is not done in isolation, input validation is still a useful technique, since it may significantly reduce your attack surface, allow you to detect some attacks, and provide other security benefits that proper encoding does not address. Ensure that you perform input validation at well-defined interfaces within the application. This will help protect the application even if a component is reused or moved elsewhere.
  10. Áp đặt quy tắc bằng chuyển đổi · Kiến trúc và thiết kếWhen the set of acceptable objects, such as filenames or URLs, is limited or known, create a mapping from a set of fixed input values (such as numeric IDs) to the actual filenames or URLs, and reject all other inputs.
  11. Tường lửa · Vận hành · Hiệu quả: KháUse an application firewall that can detect attacks against this weakness. It can be beneficial in cases in which the code cannot be fixed (because it is controlled by a third party), as an emergency prevention measure while more comprehensive software assurance measures are applied, or to provide defense in depth [REF-1481].An application firewall might not cover all possible input vectors. In addition, attack techniques might be available to bypass the protection mechanism, such as using malformed inputs that can still be processed by the component that receives those inputs. Depending on functionality, an application firewall might inadvertently reject or modify legitimate requests. Finally, some manual effort may be required for customization.
  12. Gia cố môi trường · Vận hành, Hiện thực hóaWhen using PHP, configure the application so that it does not use register_globals. During implementation, develop the application so that it does not rely on this feature, but be wary of implementing a register_globals emulation that is subject to weaknesses such as CWE-95, CWE-621, and similar issues.

Cách phát hiện trong hệ thống

Dùng phân tích tĩnh tự động, bao gồm phân tích luồng dữ liệu, để phát hiện luồng từ đầu vào không tin cậy đến đầu ra web. Phương pháp này có hiệu quả ở mức vừa phải vì không thể đạt độ chính xác và phạm vi bao phủ tuyệt đối, đặc biệt khi điểm yếu trải qua nhiều thành phần. Dùng kiểm thử hộp đen với XSS Cheat Sheet hoặc công cụ tự động tạo kiểm thử để kiểm tra nhiều biến thể tấn công; với XSS lưu trữ, việc kiểm thử phải bao gồm cả bước đưa dữ liệu vào kho lưu trữ và bước dữ liệu đó được gửi tới người dùng khác, có thể xảy ra sau một khoảng thời gian đáng kể. Các phương pháp này hữu ích nhưng không bảo đảm phát hiện đầy đủ.

Dữ liệu MITRE CWE chính thức
Phương phápCách làmHiệu quả
Phân tích tĩnh tự độngUse automated static analysis tools that target this type of weakness. Many modern techniques use data flow analysis to minimize the number of false positives. This is not a perfect solution, since 100% accuracy and coverage are not feasible, especially when multiple components are involved.Khá
Hộp đenUse the XSS Cheat Sheet [REF-714] or automated test-generation tools to help launch a wide variety of attacks against your web application. The Cheat Sheet contains many subtle XSS variations that are specifically targeted against weak XSS defenses.With Stored XSS, the indirection caused by the data store can make it more difficult to find the problem. The tester must first inject the XSS string into the data store, then find the appropriate application functionality in which the XSS string is sent to other users of the application. These are two distinct steps in which the activation of the XSS can take place minutes, hours, or days after the XSS was originally injected into the data store.Khá

Lỗ hổng điển hình

Bản ghi chính thức liệt kê các ví dụ tiêu biểu sau đây, không phải toàn bộ các lỗ hổng: CVE-2024-49038 và CVE-2024-54142 liên quan đến XSS trong chức năng AI; CVE-2021-25926 và CVE-2021-25963 liên quan đến XSS phản chiếu trong phần mềm Python; còn CVE-2021-1879 mô tả XSS phổ quát trong một hệ điều hành di động. Các ví dụ khác bao gồm XSS thông qua cookie (CVE-2014-8958), tiêu đề HTTP được tạo thủ công như Referer (CVE-2017-9764, CVE-2014-5198), PATH_INFO (CVE-2008-5770), email (CVE-2008-5734), thông báo lỗi (CVE-2008-4730), trang wiki (CVE-2008-5249), ứng dụng sổ lưu bút (CVE-2006-3568, CVE-2006-3211) và một sản phẩm bảo mật (CVE-2008-0971). Các ví dụ theo chuỗi gồm CVE-2020-3580, CVE-2008-5080, CVE-2006-4308, CVE-2007-5727 và CVE-2006-3295, trong đó một điểm yếu hoặc lỗi bảo vệ khác dẫn đến XSS.

Dữ liệu MITRE CWE chính thức

Dưới đây là các lỗ hổng tiêu biểu liên quan đến CWE-79, dựa theo mức độ ưu tiên

Nguồn (20)

CWE™ Program, operated by The MITRE Corporation. Copyright © 2006–2026, The MITRE Corporation. The MITRE Corporation hereby grants you a non-exclusive, royalty-free license to use CWE for research, development, and commercial purposes. CWE Terms of Use.

Tìm hiểu thêm

Kiểm tra chuyên sâu cùng giải pháp quản lý rủi ro Web toàn diện

Giải pháp CyStack VulnScan liên tục phát hiện tài sản, xác minh lỗ hổng và giúp đội ngũ bảo mật ưu tiên khắc phục cho toàn bộ doanh nghiệp.

Khám phá CyStack VulnScan