php清理word产生的html的函数_php函数

php清理word产生的html的函数: 发布时间：2020-08-06编辑：脚本学堂

在使用FCKedior或ueditor时，从word中粘贴过来的内容，会产生很多额外的标签，本文分享一个清理这种多出的html代码的函数，有需要的朋友参考下。

说明：
使用FCKedior或ueditor时，从word中粘贴过来的，会产生好多额外的标签。
本节分享一例代码，用于清除word产生的html代码。

例子：

<?php
/**
* 清理word产生的html
* by www.jb200.com
*/
function strip_word_html($text, $allowed_tags = '<b><i><sup><sub><em><strong><u><br>')
{
mb_regex_encoding('UTF-8');
$search = array('/&lsquo;/u', '/&rsquo;/u', '/&ldquo;/u', '/&rdquo;/u', '/&mdash;/u');
$replace = array(''', ''', '"', '"', '-');
$text = preg_replace($search, $replace, $text);
$text = html_entity_decode($text, ENT_QUOTES, 'UTF-8');
if(mb_stripos($text, '/*') !== FALSE){
$text = mb_eregi_replace('#/*.*?*/#s', '', $text, 'm');
}
$text = preg_replace(array('/<([0-9]+)/'), array('< $1'), $text);
$text = strip_tags($text, $allowed_tags);
$text = preg_replace(array('/^ss+/', '/ss+$/', '/ss+/u'), array('', '', ' '), $text);
$search = array('#<(strong|b)[^>]*>(.*?)</(strong|b)>#isu', '#<(em|i)[^>]*>(.*?)</(em|i)>#isu', '#<u[^>]*>(.*?)</u>#isu');
$replace = array('<b>$2</b>', '<i>$2</i>', '<u>$1</u>');
$text = preg_replace($search, $replace, $text);
$num_matches = preg_match_all("/<!--/u", $text, $matches);
if($num_matches){
$text = preg_replace('/<!--(.)*-->/isu', '', $text);
}
return $text;
}

上一篇：PHP get_class()函数的用法举例
下一篇：PHP函数ip2long()实现IP转换成整型时出现负数的解决方法

与 php清理word产生的html的函数有关的文章

本文标题：php清理word产生的html的函数
本页链接：http://www.jb200.com/article/13099.html

php清理word产生的html的函数

php清理word产生的html的函数

与 php清理word产生的html的函数有关的文章

浏览排行

栏目分类

栏目导航

热点文章

php清理word产生的html的函数

php清理word产生的html的函数

与 php清理word产生的html的函数 有关的文章

浏览排行

栏目分类

栏目导航

热点文章

与 php清理word产生的html的函数有关的文章